LLM EXPLORER 61,561 MODELS INDEXED

Qwen3 8B GGUF by unsloth

By unsloth · 206898 downloads

Qwen3 8B GGUF is an open-source language model by unsloth. Features: 8b LLM, VRAM: 2.3GB, Context: 40K, License: apache-2.0, Quantized, LLM Explorer Score: 0.27.

  Arxiv:2309.00071 Base model:quantized:qwen/qwen...   Base model:qwen/qwen3-8b   Conversational   En   Endpoints compatible   Gguf   Q2   Quantized   Qwen   Qwen3   Region:us   Unsloth
Model Card on HF πŸ€—: https://huggingface.co/unsloth/Qwen3-8B-GGUF 

Qwen3 8B GGUF Parameters and Internals

LLM NameQwen3 8B GGUF
Repository πŸ€—https://huggingface.co/unsloth/Qwen3-8B-GGUF 
Base Model(s)  Qwen3 8B   Qwen/Qwen3-8B
Model Size8b
Required VRAM2.3 GB
Updated2026-07-16
Maintainerunsloth
Model Typeqwen3
Model Files  16.4 GB   4.8 GB   4.6 GB   3.3 GB   3.4 GB   4.1 GB   3.8 GB   5.2 GB   5.0 GB   4.8 GB   5.8 GB   5.7 GB   6.7 GB   8.7 GB   2.4 GB   2.3 GB   3.1 GB   2.6 GB   3.4 GB   3.5 GB   4.3 GB   5.1 GB   5.9 GB   7.5 GB   10.8 GB
Supported Languagesen
GGUF QuantizationYes
Quantization Typegguf|q2|q4_k|q5_k
Model ArchitectureQwen3ForCausalLM
Licenseapache-2.0
Context Length40960
Model Max Length40960
Transformers Version4.51.3
Vocabulary Size151936
Torch Data Typebfloat16

Best Alternatives to Qwen3 8B GGUF

Best Alternatives
Context / RAM
Downloads
Likes
DeepSeek R1 0528 Qwen3 8B GGUF128K / 2.3 GB52800436
Qwen3 8B 128K GGUF128K / 2.3 GB210628
Qwen3 8B Houdini VEX V1140K / 16.4 GB5963
AOS Qwen3 8B Grpo Merged40K / 16.4 GB4850
...S Qwen3 8B Narrated Sft Merged40K / 16.4 GB6200
AOS Qwen3 8B Narrated Merged40K / 16.4 GB4130
Qwen3 8B MedReasonPath40K / 16.4 GB3491
Midas FableAgent 8B40K / 16.4 GB1100
CustomThinker 0 8B40K / 16.4 GB20529
BehChat SFTv3 Ckpt1128K / 16.4 GB60
Note: green Score (e.g. "73.2") means that the model is better than unsloth/Qwen3-8B-GGUF.