LLM EXPLORER 55,709 MODELS INDEXED

Qwen2 VL 7B Instruct 8bit by mlx-community

By mlx-community · 36 downloads

Qwen2 VL 7B Instruct 8bit is an open-source language model by mlx-community. Features: 7b LLM, VRAM: 8.9GB, Context: 32K, License: apache-2.0, Quantized, Instruction-Based.

  8bit   Conversational   En   Endpoints compatible   Image-text-to-text   Instruct   Mlx   Multimodal   Quantized   Qwen2 vl   Region:us   Safetensors   Sharded   Tensorflow

Qwen2 VL 7B Instruct 8bit Parameters and Internals

LLM NameQwen2 VL 7B Instruct 8bit
Repository πŸ€—https://huggingface.co/mlx-community/Qwen2-VL-7B-Instruct-8bit 
Model Size7b
Required VRAM8.9 GB
Updated2026-08-08
Maintainermlx-community
Model Typeqwen2_vl
Instruction-BasedYes
Model Files  5.4 GB: 1-of-2   3.5 GB: 2-of-2
Supported Languagesen
Quantization Type8bit
Model ArchitectureQwen2VLForConditionalGeneration
Licenseapache-2.0
Context Length32768
Model Max Length32768
Transformers Version4.41.2
Tokenizer ClassQwen2Tokenizer
Padding Token<|endoftext|>
Vocabulary Size152064
Torch Data Typebfloat16
Errorsreplace

Best Alternatives to Qwen2 VL 7B Instruct 8bit

Best Alternatives
Context / RAM
Downloads
Likes
Qwen2 VL 7B Instruct 4bit32K / 4.7 GB6482
OlmOCR 7B 0225 Preview 4bit32K / 5.6 GB221
...L 7B Instruct Abliterated 8bit32K / 8.9 GB291
Qwen2 VL 7B Instruct32K / 16.7 GB3337
Qwen2 VL 7B Instruct AWQ32K / 6.9 GB1075222
Qwen2 VL 7B Instruct GPTQ Int432K / 7 GB258814
Qwen2 VL 7B Instruct GPTQ Int832K / 10.2 GB126631
Qwen2 VL 7B Instruct Bf1632K / 16.6 GB504
Note: green Score (e.g. "73.2") means that the model is better than mlx-community/Qwen2-VL-7B-Instruct-8bit.