LLM EXPLORER 57,252 MODELS INDEXED

Qwen3 VL 32B Thinking by Qwen

By Qwen · 15666 downloads

Qwen3 VL 32B Thinking is an open-source language model by Qwen. Features: 32b LLM, VRAM: 67GB, License: apache-2.0.

  Arxiv:2308.12966   Arxiv:2409.12191   Arxiv:2502.13923   Arxiv:2505.09388   Conversational   Deploy:azure   Endpoints compatible   Image-text-to-text   Qwen3 vl   Region:us   Safetensors   Sharded   Tensorflow

Qwen3 VL 32B Thinking Parameters and Internals

LLM NameQwen3 VL 32B Thinking
Repository πŸ€—https://huggingface.co/Qwen/Qwen3-VL-32B-Thinking 
Model Size32b
Required VRAM67 GB
Updated2026-08-10
MaintainerQwen
Model Typeqwen3_vl
Model Files  4.9 GB: 1-of-14   4.9 GB: 2-of-14   4.9 GB: 3-of-14   4.9 GB: 4-of-14   4.9 GB: 5-of-14   4.9 GB: 6-of-14   4.9 GB: 7-of-14   4.9 GB: 8-of-14   4.9 GB: 9-of-14   4.9 GB: 10-of-14   4.9 GB: 11-of-14   4.9 GB: 12-of-14   4.9 GB: 13-of-14   3.3 GB: 14-of-14
Model ArchitectureQwen3VLForConditionalGeneration
Licenseapache-2.0
Model Max Length262144
Transformers Version4.57.0.dev0
Tokenizer ClassQwen2Tokenizer
Padding Token<|endoftext|>
Errorsreplace

Quantized Models of the Qwen3 VL 32B Thinking

Model
Likes
Downloads
VRAM
... 32B Thinking Unsloth Bnb 4bit22528 GB
Qwen3 VL 32B Thinking Bnb 4bit25520 GB

Best Alternatives to Qwen3 VL 32B Thinking

Best Alternatives
Context / RAM
Downloads
Likes
Qwen3 VL 32B Instruct0K / 67 GB2881482223
Qwen3 VL 32B Instruct FP80K / 35.7 GB38061947
Qwen3 VL 32B Thinking FP80K / 35.7 GB1143127
Qwen3 VL 32B Thinking FP80K / 35.7 GB272
Qwen3 VL 32B Instruct FP80K / 35.7 GB301
Cosmos Reason2 32B0K / 66.4 GB266214
Qwen3 VL 32B Instruct0K / 67 GB3305
Qwen3 VL 32B Thinking0K / 67 GB323
Qwen3 VL 32B Instruct Bnb 4bit0K / 20.5 GB2988466
...truct Ultra Uncensored Heretic0K / 66.7 GB31011
Note: green Score (e.g. "73.2") means that the model is better than Qwen/Qwen3-VL-32B-Thinking.