LLM EXPLORER 59,420 MODELS INDEXED

Qwen3 14B 128K GGUF by unsloth

By unsloth · 2431 downloads

Qwen3 14B 128K GGUF is an open-source language model by unsloth. Features: 14b LLM, VRAM: 3.8GB, Context: 128K, License: apache-2.0, Quantized, LLM Explorer Score: 0.2.

  Arxiv:2309.00071 Base model:quantized:qwen/qwen...   Base model:qwen/qwen3-14b   Conversational   En   Endpoints compatible   Gguf   Q2   Quantized   Qwen   Qwen3   Region:us   Unsloth

Qwen3 14B 128K GGUF Parameters and Internals

LLM NameQwen3 14B 128K GGUF
Repository πŸ€—https://huggingface.co/unsloth/Qwen3-14B-128K-GGUF 
Base Model(s)  Qwen3 14B   Qwen/Qwen3-14B
Model Size14b
Required VRAM3.8 GB
Updated2026-07-22
Maintainerunsloth
Model Typeqwen3
Model Files  29.5 GB   8.5 GB   8.1 GB   5.8 GB   5.9 GB   7.3 GB   6.7 GB   8.5 GB   9.4 GB   9.0 GB   8.6 GB   10.5 GB   10.3 GB   12.1 GB   15.7 GB   4.1 GB   3.8 GB   5.4 GB   4.5 GB   6.0 GB   6.1 GB   7.6 GB   9.2 GB   10.5 GB   13.3 GB   18.8 GB
Supported Languagesen
GGUF QuantizationYes
Quantization Typegguf|q2|q4_k|q5_k
Model ArchitectureQwen3ForCausalLM
Licenseapache-2.0
Context Length131072
Model Max Length131072
Transformers Version4.51.3
Vocabulary Size151936
Torch Data Typebfloat16

Best Alternatives to Qwen3 14B 128K GGUF

Best Alternatives
Context / RAM
Downloads
Likes
Qwen3 14B GGUF40K / 3.8 GB32233133
Qwen3 14B Unsloth Bnb 4bit40K / 11.2 GB18052117
Prototie Ai40K / 29.5 GB3850
Qwen3 14B 4bit40K / 8.3 GB808375
Qwen3 14B MLX 4bit40K / 8.3 GB495366
Qwen3 14B MLX 8bit40K / 15.2 GB164565
Qwen3 14B Bnb 4bit40K / 9.9 GB85126
Hermes 4 14B 4bit40K / 8.3 GB11225
Hermes 4 14B 8bit40K / 15.7 GB3682
Qwen3 14B MLX 4bit40K / 7.9 GB132814
Note: green Score (e.g. "73.2") means that the model is better than unsloth/Qwen3-14B-128K-GGUF.