LLM EXPLORER 56,604 MODELS INDEXED

Qwen3 30B A3B GGUF by unsloth

By unsloth · 54806 downloads

Qwen3 30B A3B GGUF is an open-source language model by unsloth. Features: 30b LLM, VRAM: 9GB, Context: 40K, License: apache-2.0, Quantized, LLM Explorer Score: 0.26.

  Arxiv:2309.00071 Base model:quantized:qwen/qwen...   Base model:qwen/qwen3-30b-a3b   Conversational   En   Endpoints compatible   Gguf   Imatrix   Q2   Quantized   Qwen   Qwen3   Qwen3 moe   Region:us   Unsloth

Qwen3 30B A3B GGUF Parameters and Internals

LLM NameQwen3 30B A3B GGUF
Repository πŸ€—https://huggingface.co/unsloth/Qwen3-30B-A3B-GGUF 
Base Model(s)  Qwen3 30B A3B   Qwen/Qwen3-30B-A3B
Model Size30b
Required VRAM9 GB
Updated2026-07-23
Maintainerunsloth
Model Typeqwen3_moe
Model Files  17.3 GB   16.4 GB   11.3 GB   11.3 GB   14.7 GB   13.3 GB   17.4 GB   19.2 GB   18.6 GB   17.5 GB   21.7 GB   21.1 GB   25.1 GB   32.5 GB   9.7 GB   9.0 GB   10.9 GB   10.4 GB   12.9 GB   11.8 GB   13.8 GB   17.7 GB   21.7 GB   26.3 GB   36.0 GB
Supported Languagesen
GGUF QuantizationYes
Quantization Typegguf|q2|q4_k|q5_k
Model ArchitectureQwen3MoeForCausalLM
Licenseapache-2.0
Context Length40960
Model Max Length40960
Transformers Version4.51.3
Vocabulary Size151936
Torch Data Typebfloat16

Best Alternatives to Qwen3 30B A3B GGUF

Best Alternatives
Context / RAM
Downloads
Likes
Qwen3 30B A3B 128K GGUF128K / 9 GB273061
...en3 30B A3B Instruct 2507 4bit256K / 17.2 GB8376612
...30B A3B Instruct 2507 MLX 4bit256K / 17.2 GB246528
...30B A3B Instruct 2507 MLX 8bit256K / 32.5 GB246553
...n3 Coder 30B A3B Instruct 4bit256K / 17.2 GB725233
...r 30B A3B Instruct 4bit Dwq V2256K / 17.2 GB134210
...n3 Coder 30B A3B Instruct 8bit256K / 32.5 GB12777
...oder 30B A3B Instruct 4bit DWQ256K / 17.2 GB6445
...30B A3B Thinking 2507 MLX 4bit256K / 17.2 GB5761
...en3 30B A3B Thinking 2507 4bit256K / 17.2 GB2324
Note: green Score (e.g. "73.2") means that the model is better than unsloth/Qwen3-30B-A3B-GGUF.