LLM EXPLORER 57,918 MODELS INDEXED

Qwen3 32B FP8 by unsloth

By unsloth · 269 downloads

Qwen3 32B FP8 is an open-source language model by unsloth. Features: 32b LLM, VRAM: 34.5GB, Context: 40K, License: apache-2.0, LLM Explorer Score: 0.18.

  Arxiv:2309.00071 Base model:quantized:qwen/qwen...   Base model:qwen/qwen3-32b-fp8   Conversational   Endpoints compatible   Fp8   Qwen3   Region:us   Safetensors   Sharded   Tensorflow   Unsloth
Model Card on HF πŸ€—: https://huggingface.co/unsloth/Qwen3-32B-FP8 

Qwen3 32B FP8 Parameters and Internals

LLM NameQwen3 32B FP8
Repository πŸ€—https://huggingface.co/unsloth/Qwen3-32B-FP8 
Base Model(s)  Qwen3 32B FP8   Qwen/Qwen3-32B-FP8
Model Size32b
Required VRAM34.5 GB
Updated2026-08-10
Maintainerunsloth
Model Typeqwen3
Model Files  5.0 GB: 1-of-7   5.0 GB: 2-of-7   4.9 GB: 3-of-7   4.9 GB: 4-of-7   4.9 GB: 5-of-7   4.9 GB: 6-of-7   4.9 GB: 7-of-7
Model ArchitectureQwen3ForCausalLM
Licenseapache-2.0
Context Length40960
Model Max Length40960
Transformers Version4.52.0.dev0
Tokenizer ClassQwen2Tokenizer
Padding Token<|vision_pad|>
Vocabulary Size151936
Torch Data Typebfloat16
Errorsreplace

Best Alternatives to Qwen3 32B FP8

Best Alternatives
Context / RAM
Downloads
Likes
...ll Qwen3 32B Preview7 QAT 200K195K / 65.8 GB103
...ll Qwen3 32B Preview6 QAT 200K195K / 65.8 GB62
...ll Qwen3 32B Preview4 QAT 200K195K / 65.8 GB132
Qwen3 32B AWorld128K / 65.8 GB3715
Qwen3 32B40K / 65.8 GB541115
Qwen3 32B40K / 65.6 GB6155624715
OpenThinkerAgent 32B40K / 65.8 GB90229
Qwen3 Swallow 32B SFT V0.240K / 65.6 GB244110
...gy GPT Regulatorio 32B Safe V240K / 131.6 GB2050
Qwen3 Swallow 32B RL V0.240K / 65.1 GB103473
Note: green Score (e.g. "73.2") means that the model is better than unsloth/Qwen3-32B-FP8.