LLM EXPLORER 59,420 MODELS INDEXED

Qwen3 VL 4B Thinking 3bit by mlx-community

By mlx-community · 23 downloads

Qwen3 VL 4B Thinking 3bit is an open-source language model by mlx-community. Features: 4b LLM, VRAM: 2.6GB, License: apache-2.0, Quantized, LLM Explorer Score: 0.2.

  3-bit   3bit   Conversational   Image-text-to-text   Mlx   Quantized   Qwen3 vl   Region:us   Safetensors

Qwen3 VL 4B Thinking 3bit Parameters and Internals

LLM NameQwen3 VL 4B Thinking 3bit
Repository πŸ€—https://huggingface.co/mlx-community/Qwen3-VL-4B-Thinking-3bit 
Model Size4b
Required VRAM2.6 GB
Updated2026-08-24
Maintainermlx-community
Model Typeqwen3_vl
Model Files  2.6 GB
Quantization Type3bit
Model ArchitectureQwen3VLForConditionalGeneration
Licenseapache-2.0
Model Max Length262144
Transformers Version4.57.0
Padding Token<|endoftext|>
Errorsreplace

Best Alternatives to Qwen3 VL 4B Thinking 3bit

Best Alternatives
Context / RAM
Downloads
Likes
...L 4B Instruct Unsloth Bnb 4bit0K / 4.3 GB25381712
Qwen3 VL 4B Instruct 4bit0K / 3.1 GB174167
Qwen3 VL 4B Instruct Bnb 4bit0K / 3.5 GB14923
...L 4B Thinking Unsloth Bnb 4bit0K / 4.9 GB12352
Qwen3 VL 4B Instruct 8bit0K / 5.1 GB5873
Qwen3 VL 4B Instruct 3bit0K / 2.6 GB2754
Qwen3 VL 4B Thinking Bnb 4bit0K / 3.5 GB581
Qwen3 VL 4B Instruct0K / 8.9 GB3611570437
Qwen3 VL 4B Instruct FP80K / 6.1 GB30110365
Qwen3 VL 4B Instruct AWQ 4bit0K / 4.4 GB2446739
Note: green Score (e.g. "73.2") means that the model is better than mlx-community/Qwen3-VL-4B-Thinking-3bit.