LLM EXPLORER 61,918 MODELS INDEXED

QVQ 72B Preview 8bit by mlx-community

By mlx-community · 22 downloads

QVQ 72B Preview 8bit is an open-source language model by mlx-community. Features: 72b LLM, VRAM: 77.8GB, Context: 125K, License: other, Quantized, LLM Explorer Score: 0.16.

  8bit Base model:finetune:qwen/qwen2...   Base model:qwen/qwen2-vl-72b   Chat   Conversational   En   Endpoints compatible   Image-text-to-text   Mlx   Quantized   Qwen2 vl   Region:us   Safetensors   Sharded   Tensorflow

QVQ 72B Preview 8bit Parameters and Internals

LLM NameQVQ 72B Preview 8bit
Repository πŸ€—https://huggingface.co/mlx-community/QVQ-72B-Preview-8bit 
Base Model(s)  Qwen2 VL 72B   Qwen/Qwen2-VL-72B
Model Size72b
Required VRAM77.8 GB
Updated2026-08-08
Maintainermlx-community
Model Typeqwen2_vl
Model Files  5.3 GB: 1-of-15   5.3 GB: 2-of-15   5.2 GB: 3-of-15   5.3 GB: 4-of-15   5.3 GB: 5-of-15   5.2 GB: 6-of-15   5.3 GB: 7-of-15   5.3 GB: 8-of-15   5.2 GB: 9-of-15   5.3 GB: 10-of-15   5.3 GB: 11-of-15   5.2 GB: 12-of-15   5.3 GB: 13-of-15   5.3 GB: 14-of-15   4.0 GB: 15-of-15
Supported Languagesen
Quantization Type8bit
Model ArchitectureQwen2VLForConditionalGeneration
Licenseother
Context Length128000
Model Max Length128000
Transformers Version4.41.2
Tokenizer ClassQwen2Tokenizer
Padding Token<|endoftext|>
Vocabulary Size152064
Torch Data Typebfloat16
Errorsreplace

Best Alternatives to QVQ 72B Preview 8bit

Best Alternatives
Context / RAM
Downloads
Likes
QVQ 72B Preview 4bit125K / 41.3 GB207
QVQ 72B Preview 6bit125K / 59.7 GB342
QVQ 72B Preview Bnb 4bit125K / 41.8 GB155
QVQ 72B Preview 3bit125K / 32.2 GB215
Qwen2 VL 72B Instruct Bnb 4bit32K / 41.8 GB3955
Qwen2 VL 72B Instruct 8bit32K / 77.8 GB794
Qwen2 VL 72B Instruct 4bit32K / 41.3 GB725
QVQ 72B Preview125K / 146.8 GB8572610
QVQ 72B Preview Bf16125K / 147.7 GB233
Qwen2 VL 72B Instruct32K / 146.8 GB60157311
Note: green Score (e.g. "73.2") means that the model is better than mlx-community/QVQ-72B-Preview-8bit.