LLM EXPLORER 55,709 MODELS INDEXED

QVQ 72B Preview 3bit by mlx-community

By mlx-community · 21 downloads

QVQ 72B Preview 3bit is an open-source language model by mlx-community. Features: 72b LLM, VRAM: 32.2GB, Context: 125K, License: other, Quantized.

  3bit Base model:finetune:qwen/qwen2...   Base model:qwen/qwen2-vl-72b   Chat   Conversational   En   Endpoints compatible   Image-text-to-text   Mlx   Quantized   Qwen2 vl   Region:us   Safetensors   Sharded   Tensorflow

QVQ 72B Preview 3bit Parameters and Internals

LLM NameQVQ 72B Preview 3bit
Repository πŸ€—https://huggingface.co/mlx-community/QVQ-72B-Preview-3bit 
Base Model(s)  Qwen/Qwen2-VL-72B   Qwen/Qwen2-VL-72B
Model Size72b
Required VRAM32.2 GB
Updated2026-08-08
Maintainermlx-community
Model Typeqwen2_vl
Model Files  5.4 GB: 1-of-7   5.4 GB: 2-of-7   5.3 GB: 3-of-7   5.4 GB: 4-of-7   5.3 GB: 5-of-7   4.9 GB: 6-of-7   0.5 GB: 7-of-7
Supported Languagesen
Quantization Type3bit
Model ArchitectureQwen2VLForConditionalGeneration
Licenseother
Context Length128000
Model Max Length128000
Transformers Version4.41.2
Tokenizer ClassQwen2Tokenizer
Padding Token<|endoftext|>
Vocabulary Size152064
Torch Data Typebfloat16
Errorsreplace

Best Alternatives to QVQ 72B Preview 3bit

Best Alternatives
Context / RAM
Downloads
Likes
QVQ 72B Preview 4bit125K / 41.3 GB207
QVQ 72B Preview 6bit125K / 59.7 GB342
QVQ 72B Preview 8bit125K / 77.8 GB223
Qwen2 VL 72B Instruct 4bit32K / 41.3 GB725
Qwen2 VL 72B Instruct 8bit32K / 77.8 GB794
QVQ 72B Preview Bf16125K / 147.7 GB233
UI TARS 72B DPO Bf1632K / 147.7 GB301
UI TARS 72B SFT Bf1632K / 147.7 GB321