LLM EXPLORER 60,384 MODELS INDEXED

Qwen3 30B A3B GPTQ Int4 by Qwen

By Qwen · 169347 downloads

Qwen3 30B A3B GPTQ Int4 is an open-source language model by Qwen. Features: 30b LLM, VRAM: 16.9GB, Context: 40K, License: apache-2.0, Quantized, LLM Explorer Score: 0.26.

  Arxiv:2309.00071   Arxiv:2505.09388   4-bit Base model:quantized:qwen/qwen...   Base model:qwen/qwen3-30b-a3b   Conversational   Endpoints compatible   Gptq   Quantized   Qwen3 moe   Region:us   Safetensors

Qwen3 30B A3B GPTQ Int4 Parameters and Internals

LLM NameQwen3 30B A3B GPTQ Int4
Repository πŸ€—https://huggingface.co/Qwen/Qwen3-30B-A3B-GPTQ-Int4 
Base Model(s)  Qwen3 30B A3B   Qwen/Qwen3-30B-A3B
Model Size30b
Required VRAM16.9 GB
Updated2026-08-04
MaintainerQwen
Model Typeqwen3_moe
Model Files  16.9 GB
GPTQ QuantizationYes
Quantization Typegptq
Model ArchitectureQwen3MoeForCausalLM
Licenseapache-2.0
Context Length40960
Model Max Length40960
Transformers Version4.51.3
Tokenizer ClassQwen2Tokenizer
Padding Token<|endoftext|>
Vocabulary Size151936
Torch Data Typefloat16
Errorsreplace

Best Alternatives to Qwen3 30B A3B GPTQ Int4

Best Alternatives
Context / RAM
Downloads
Likes
...der 30B A3B Instruct GPTQ 4bit256K / 16.7 GB10555
Qwen3 30B A3B GPTQ40K / 16.9 GB253
...en3 30B A3B Instruct 2507 4bit256K / 17.2 GB8376612
...30B A3B Instruct 2507 MLX 4bit256K / 17.2 GB246528
...30B A3B Instruct 2507 MLX 8bit256K / 32.5 GB246553
...n3 Coder 30B A3B Instruct 4bit256K / 17.2 GB725233
Unslopper 30B A3B 6bit256K / 24.7 GB721
...r 30B A3B Instruct 4bit Dwq V2256K / 17.2 GB134210
...n3 Coder 30B A3B Instruct 8bit256K / 32.5 GB12777
...oder 30B A3B Instruct 4bit DWQ256K / 17.2 GB6445
Note: green Score (e.g. "73.2") means that the model is better than Qwen/Qwen3-30B-A3B-GPTQ-Int4.