LLM EXPLORER 60,384 MODELS INDEXED

Qwen3 30B A3B GPTQ by AlphaGaO

By AlphaGaO · 25 downloads

Qwen3 30B A3B GPTQ is an open-source language model by AlphaGaO. Features: 30b LLM, VRAM: 16.9GB, Context: 40K, License: apache-2.0, Quantized, LLM Explorer Score: 0.18.

  Arxiv:2309.00071   4-bit Base model:quantized:qwen/qwen... Base model:qwen/qwen3-30b-a3b-...   Conversational   Endpoints compatible   Gptq   Quantized   Qwen3 moe   Region:us   Safetensors   Sharded   Tensorflow

Qwen3 30B A3B GPTQ Parameters and Internals

LLM NameQwen3 30B A3B GPTQ
Repository πŸ€—https://huggingface.co/AlphaGaO/Qwen3-30B-A3B-GPTQ 
Base Model(s)  Qwen3 30B A3B Base   Qwen/Qwen3-30B-A3B-Base
Model Size30b
Required VRAM16.9 GB
Updated2026-09-04
MaintainerAlphaGaO
Model Typeqwen3_moe
Model Files  4.0 GB: 1-of-5   4.0 GB: 2-of-5   4.0 GB: 3-of-5   4.0 GB: 4-of-5   0.9 GB: 5-of-5
GPTQ QuantizationYes
Quantization Typegptq
Model ArchitectureQwen3MoeForCausalLM
Licenseapache-2.0
Context Length40960
Model Max Length40960
Transformers Version4.51.3
Tokenizer ClassQwen2TokenizerFast
Padding Token<unk>
Vocabulary Size151936
Torch Data Typebfloat16
Errorsreplace

Best Alternatives to Qwen3 30B A3B GPTQ

Best Alternatives
Context / RAM
Downloads
Likes
...der 30B A3B Instruct GPTQ 4bit256K / 16.7 GB10555
Qwen3 30B A3B GPTQ Int440K / 16.9 GB16934755
...en3 30B A3B Instruct 2507 4bit256K / 17.2 GB8376612
...30B A3B Instruct 2507 MLX 4bit256K / 17.2 GB246528
...30B A3B Instruct 2507 MLX 8bit256K / 32.5 GB246553
...n3 Coder 30B A3B Instruct 4bit256K / 17.2 GB725233
Unslopper 30B A3B 6bit256K / 24.7 GB721
...r 30B A3B Instruct 4bit Dwq V2256K / 17.2 GB134210
...n3 Coder 30B A3B Instruct 8bit256K / 32.5 GB12777
...oder 30B A3B Instruct 4bit DWQ256K / 17.2 GB6445
Note: green Score (e.g. "73.2") means that the model is better than AlphaGaO/Qwen3-30B-A3B-GPTQ.