LLM EXPLORER 57,252 MODELS INDEXED

Qwen3 VL 2B Thinking 4bit by mlx-community

By mlx-community · 40 downloads

Qwen3 VL 2B Thinking 4bit is an open-source language model by mlx-community. Features: 2b LLM, VRAM: 1.8GB, License: apache-2.0, Quantized, LLM Explorer Score: 0.21.

  4-bit   4bit   Conversational   Endpoints compatible   Image-text-to-text   Mlx   Quantized   Qwen3 vl   Region:us   Safetensors

Qwen3 VL 2B Thinking 4bit Parameters and Internals

LLM NameQwen3 VL 2B Thinking 4bit
Repository πŸ€—https://huggingface.co/mlx-community/Qwen3-VL-2B-Thinking-4bit 
Model Size2b
Required VRAM1.8 GB
Updated2026-08-08
Maintainermlx-community
Model Typeqwen3_vl
Model Files  1.8 GB
Quantization Type4bit
Model ArchitectureQwen3VLForConditionalGeneration
Licenseapache-2.0
Model Max Length262144
Transformers Version4.57.1
Padding Token<|endoftext|>
Errorsreplace

Best Alternatives to Qwen3 VL 2B Thinking 4bit

Best Alternatives
Context / RAM
Downloads
Likes
...L 2B Instruct Unsloth Bnb 4bit0K / 2.4 GB1648448
...i Qwen3 VL 2B Instruct Ab 4bit0K / 1.8 GB1441
Qwen3 VL 2B Instruct 4bit0K / 1.8 GB18812
Qwen3 VL 2B Instruct 8bit0K / 2.6 GB761
Qwen3 VL 2B Instruct 6bit0K / 2.2 GB361
Qwen3 VL 2B Instruct Bnb 4bit0K / 2.2 GB9383
...L 2B Thinking Unsloth Bnb 4bit0K / 2.3 GB1771
Qwen 3 VL 2B Instruct Heretic0K / 4.3 GB442
Typhoon Ocr1.5 2B0K / 4.3 GB5502925
Qwen3 VL 2B Instruct FP80K / 3.5 GB6506343
Note: green Score (e.g. "73.2") means that the model is better than mlx-community/Qwen3-VL-2B-Thinking-4bit.