LLM EXPLORER 55,709 MODELS INDEXED

Qwen2 VL 2B 8bit by mlx-community

By mlx-community · 21 downloads

Qwen2 VL 2B 8bit is an open-source language model by mlx-community. Features: 2b LLM, VRAM: 2.4GB, Context: 32K, License: apache-2.0, Quantized.

  8bit   Conversational   En   Endpoints compatible   Image-text-to-text   Mlx   Multimodal   Quantized   Qwen2 vl   Region:us   Safetensors

Qwen2 VL 2B 8bit Parameters and Internals

LLM NameQwen2 VL 2B 8bit
Repository πŸ€—https://huggingface.co/mlx-community/Qwen2-VL-2B-8bit 
Model Size2b
Required VRAM2.4 GB
Updated2026-08-08
Maintainermlx-community
Model Typeqwen2_vl
Model Files  2.4 GB
Supported Languagesen
Quantization Type8bit
Model ArchitectureQwen2VLForConditionalGeneration
Licenseapache-2.0
Context Length32768
Model Max Length32768
Transformers Version4.41.2
Tokenizer ClassQwen2Tokenizer
Padding Token<|endoftext|>
Vocabulary Size151936
Torch Data Typebfloat16
Errorsreplace

Best Alternatives to Qwen2 VL 2B 8bit

Best Alternatives
Context / RAM
Downloads
Likes
Qwen2 VL 2B 4bit32K / 1.2 GB343
Qwen2 VL 2B Instruct 4bit32K / 1.2 GB38997
Qwen2 VL 2B Instruct 8bit32K / 2.4 GB992
ShowUI 2B 6bit V232K / 2.8 GB221
Qwen2 VL 2B Instruct AWQ32K / 3 GB375226
Qwen2 VL 2B Instruct GPTQ Int432K / 2.5 GB79328
Qwen2 VL 2B Instruct GPTQ Int832K / 3.1 GB47117
Qwen2 VL 2B Instruct32K / 4.4 GB324
UGround V1 2B32K / 4.4 GB134211
ShowUI 2B32K / 4.4 GB3756279
Note: green Score (e.g. "73.2") means that the model is better than mlx-community/Qwen2-VL-2B-8bit.