LLM EXPLORER 57,918 MODELS INDEXED

QVQ 72B Preview Bnb 4bit by unsloth

By unsloth · 15 downloads

QVQ 72B Preview Bnb 4bit is an open-source language model by unsloth. Features: 72b LLM, VRAM: 41.8GB, Context: 125K, License: other, Quantized, LLM Explorer Score: 0.16.

  Arxiv:2409.12191   4-bit   4bit Base model:quantized:qwen/qvq-... Base model:qwen/qvq-72b-previe...   Bitsandbytes   Chat   Conversational   Deploy:azure   En   Endpoints compatible   Image-text-to-text   Quantized   Qwen   Qwen2 vl   Region:us   Safetensors   Sharded   Tensorflow

QVQ 72B Preview Bnb 4bit Parameters and Internals

LLM NameQVQ 72B Preview Bnb 4bit
Repository πŸ€—https://huggingface.co/unsloth/QVQ-72B-Preview-bnb-4bit 
Base Model(s)  QVQ 72B Preview   Qwen/QVQ-72B-Preview
Model Size72b
Required VRAM41.8 GB
Updated2026-08-10
Maintainerunsloth
Model Typeqwen2_vl
Model Files  5.0 GB: 1-of-9   5.0 GB: 2-of-9   5.0 GB: 3-of-9   5.0 GB: 4-of-9   5.0 GB: 5-of-9   5.0 GB: 6-of-9   5.0 GB: 7-of-9   4.3 GB: 8-of-9   2.5 GB: 9-of-9
Supported Languagesen
Quantization Type4bit
Model ArchitectureQwen2VLForConditionalGeneration
Licenseother
Context Length128000
Model Max Length128000
Transformers Version4.47.1
Tokenizer ClassQwen2Tokenizer
Padding Token<|vision_pad|>
Vocabulary Size152064
Torch Data Typebfloat16
Errorsreplace

Best Alternatives to QVQ 72B Preview Bnb 4bit

Best Alternatives
Context / RAM
Downloads
Likes
QVQ 72B Preview 4bit125K / 41.3 GB207
QVQ 72B Preview 6bit125K / 59.7 GB342
QVQ 72B Preview 3bit125K / 32.2 GB215
QVQ 72B Preview 8bit125K / 77.8 GB223
Qwen2 VL 72B Instruct Bnb 4bit32K / 41.8 GB3955
Qwen2 VL 72B Instruct 8bit32K / 77.8 GB794
Qwen2 VL 72B Instruct 4bit32K / 41.3 GB725
QVQ 72B Preview125K / 146.8 GB8572610
QVQ 72B Preview Bf16125K / 147.7 GB233
Qwen2 VL 72B Instruct32K / 146.8 GB60157311
Note: green Score (e.g. "73.2") means that the model is better than unsloth/QVQ-72B-Preview-bnb-4bit.