LLM EXPLORER 60,702 MODELS INDEXED

Qwen3 32B Bnb 4bit by unsloth

By unsloth · 75179 downloads

Qwen3 32B Bnb 4bit is an open-source language model by unsloth. Features: 32b LLM, VRAM: 19.2GB, Context: 40K, Quantized, LLM Explorer Score: 0.25.

  Arxiv:2309.00071   4-bit   4bit Base model:quantized:qwen/qwen...   Base model:qwen/qwen3-32b   Bitsandbytes   Quantized   Qwen3   Region:us   Safetensors   Sharded   Tensorflow   Unsloth

Qwen3 32B Bnb 4bit Parameters and Internals

LLM NameQwen3 32B Bnb 4bit
Repository πŸ€—https://huggingface.co/unsloth/Qwen3-32B-bnb-4bit 
Base Model(s)  Qwen3 32B   Qwen/Qwen3-32B
Model Size32b
Required VRAM19.2 GB
Updated2026-09-08
Maintainerunsloth
Model Typeqwen3
Model Files  4.9 GB: 1-of-4   5.0 GB: 2-of-4   5.0 GB: 3-of-4   4.3 GB: 4-of-4
Quantization Type4bit
Model ArchitectureQwen3ForCausalLM
Context Length40960
Model Max Length40960
Transformers Version4.51.3
Tokenizer ClassQwen2Tokenizer
Padding Token<|vision_pad|>
Vocabulary Size151936
Torch Data Typebfloat16
Errorsreplace

Quantized Models of the Qwen3 32B Bnb 4bit

Model
Likes
Downloads
VRAM
...llcpt E5 Baseline Merged 16bit02065 GB

Best Alternatives to Qwen3 32B Bnb 4bit

Best Alternatives
Context / RAM
Downloads
Likes
... E5 Grad Accum 16 Merged 16bit40K / 65.8 GB160
...llcpt E5 Baseline Merged 16bit40K / 65.8 GB200
...line Merged 16bit Merged 16bit40K / 65.8 GB130
SERA 32B 6bit40K / 26.6 GB272
SERA 32B 4bit40K / 18.5 GB121
SERA 32B GA 4bit40K / 18.5 GB71
Qwen3 32B Unsloth Bnb 4bit40K / 39.5 GB667415
Qwen3 32B 4bit40K / 18.5 GB24366
Qwen3 32B MLX 8bit40K / 33.7 GB74912
Qwen3 32B MLX 4bit40K / 17.4 GB5738