LLM EXPLORER 57,252 MODELS INDEXED

Qwen3 VL 4B Thinking by Qwen

By Qwen · 21715 downloads

Qwen3 VL 4B Thinking is an open-source language model by Qwen. Features: 4b LLM, VRAM: 8.9GB, License: apache-2.0.

  Arxiv:2308.12966   Arxiv:2409.12191   Arxiv:2502.13923   Arxiv:2505.09388   Conversational   Deploy:azure   Endpoints compatible   Image-text-to-text   Qwen3 vl   Region:us   Safetensors   Sharded   Tensorflow

Qwen3 VL 4B Thinking Parameters and Internals

LLM NameQwen3 VL 4B Thinking
Repository πŸ€—https://huggingface.co/Qwen/Qwen3-VL-4B-Thinking 
Model Size4b
Required VRAM8.9 GB
Updated2026-08-10
MaintainerQwen
Model Typeqwen3_vl
Model Files  5.0 GB: 1-of-2   3.9 GB: 2-of-2
Model ArchitectureQwen3VLForConditionalGeneration
Licenseapache-2.0
Model Max Length262144
Transformers Version4.57.0.dev0
Tokenizer ClassQwen2Tokenizer
Padding Token<|endoftext|>
Errorsreplace

Quantized Models of the Qwen3 VL 4B Thinking

Model
Likes
Downloads
VRAM
...L 4B Thinking Unsloth Bnb 4bit212354 GB
Qwen3 VL 4B Thinking Bnb 4bit1583 GB

Best Alternatives to Qwen3 VL 4B Thinking

Best Alternatives
Context / RAM
Downloads
Likes
Qwen3 VL 4B Instruct0K / 8.9 GB3611570437
Qwen3 VL 4B Instruct FP80K / 6.1 GB30110365
OMEGA 4B SpatialThink 08040K / 8.9 GB161
GELab Zero 4B Preview0K / 8.9 GB235154
ZwZ 4B0K / 9.7 GB15532
Spatial SSRL Qwen3VL 4B0K / 9.7 GB10114
Qwen3 VL 4B Thinking Bf160K / 8.8 GB301
Qwen3 VL 4B Thinking FP80K / 6.1 GB272330
Qwen3 VL 4B Thinking FP80K / 6.1 GB2041
Qwen3 VL 4B Instruct FP80K / 6.1 GB2673
Note: green Score (e.g. "73.2") means that the model is better than Qwen/Qwen3-VL-4B-Thinking.