LLM EXPLORER 57,252 MODELS INDEXED

Qwen3 30B A3B 128K GGUF by unsloth

By unsloth · 2730 downloads

Qwen3 30B A3B 128K GGUF is an open-source language model by unsloth. Features: 30b LLM, VRAM: 9GB, Context: 128K, License: apache-2.0, Quantized, LLM Explorer Score: 0.21.

  Arxiv:2309.00071 Base model:quantized:qwen/qwen...   Base model:qwen/qwen3-30b-a3b   Conversational   En   Endpoints compatible   Gguf   Imatrix   Q2   Quantized   Qwen   Qwen3   Qwen3 moe   Region:us   Unsloth

Qwen3 30B A3B 128K GGUF Parameters and Internals

LLM NameQwen3 30B A3B 128K GGUF
Repository πŸ€—https://huggingface.co/unsloth/Qwen3-30B-A3B-128K-GGUF 
Base Model(s)  Qwen3 30B A3B   Qwen/Qwen3-30B-A3B
Model Size30b
Required VRAM9 GB
Updated2026-08-04
Maintainerunsloth
Model Typeqwen3_moe
Model Files  17.3 GB   16.4 GB   11.3 GB   11.3 GB   14.7 GB   13.3 GB   17.4 GB   19.2 GB   18.6 GB   17.5 GB   21.7 GB   21.1 GB   25.1 GB   32.5 GB   9.7 GB   9.0 GB   10.9 GB   10.4 GB   12.9 GB   11.8 GB   13.8 GB   17.7 GB   21.7 GB   26.3 GB   36.0 GB
Supported Languagesen
GGUF QuantizationYes
Quantization Typegguf|q2|q4_k|q5_k
Model ArchitectureQwen3MoeForCausalLM
Licenseapache-2.0
Context Length131072
Model Max Length131072
Transformers Version4.51.3
Vocabulary Size151936
Torch Data Typebfloat16

Best Alternatives to Qwen3 30B A3B 128K GGUF

Best Alternatives
Context / RAM
Downloads
Likes
Qwen3 30B A3B GGUF40K / 9 GB54806284
...en3 30B A3B Instruct 2507 4bit256K / 17.2 GB8376612
...30B A3B Instruct 2507 MLX 4bit256K / 17.2 GB246528
...30B A3B Instruct 2507 MLX 8bit256K / 32.5 GB246553
...n3 Coder 30B A3B Instruct 4bit256K / 17.2 GB725233
...r 30B A3B Instruct 4bit Dwq V2256K / 17.2 GB134210
...n3 Coder 30B A3B Instruct 8bit256K / 32.5 GB12777
...oder 30B A3B Instruct 4bit DWQ256K / 17.2 GB6445
...30B A3B Thinking 2507 MLX 4bit256K / 17.2 GB5761
...en3 30B A3B Thinking 2507 4bit256K / 17.2 GB2324
Note: green Score (e.g. "73.2") means that the model is better than unsloth/Qwen3-30B-A3B-128K-GGUF.