LLM EXPLORER 57,252 MODELS INDEXED

Qwen2 VL 72B Instruct GPTQ Int8 by Qwen

By Qwen · 126 downloads

Qwen2 VL 72B Instruct GPTQ Int8 is an open-source language model by Qwen. Features: 72b LLM, VRAM: 78.4GB, Context: 32K, License: other, Quantized, Instruction-Based.

  Arxiv:2308.12966   8-bit Base model:quantized:qwen/qwen... Base model:qwen/qwen2-vl-72b-i...   Conversational   En   Gptq   Image-text-to-text   Instruct   Multimodal   Quantized   Qwen2 vl   Region:us   Safetensors   Sharded   Tensorflow

Qwen2 VL 72B Instruct GPTQ Int8 Parameters and Internals

LLM NameQwen2 VL 72B Instruct GPTQ Int8
Repository πŸ€—https://huggingface.co/Qwen/Qwen2-VL-72B-Instruct-GPTQ-Int8 
Base Model(s)  Qwen2 VL 72B Instruct   Qwen/Qwen2-VL-72B-Instruct
Model Size72b
Required VRAM78.4 GB
Updated2026-08-10
MaintainerQwen
Model Typeqwen2_vl
Instruction-BasedYes
Model Files  4.0 GB: 1-of-21   3.9 GB: 2-of-21   3.9 GB: 3-of-21   3.9 GB: 4-of-21   3.9 GB: 5-of-21   3.9 GB: 6-of-21   3.9 GB: 7-of-21   3.9 GB: 8-of-21   3.9 GB: 9-of-21   3.9 GB: 10-of-21   3.9 GB: 11-of-21   3.9 GB: 12-of-21   3.9 GB: 13-of-21   3.9 GB: 14-of-21   3.9 GB: 15-of-21   3.9 GB: 16-of-21   3.9 GB: 17-of-21   3.9 GB: 18-of-21   3.9 GB: 19-of-21   1.7 GB: 20-of-21   2.5 GB: 21-of-21
Supported Languagesen
GPTQ QuantizationYes
Quantization Typegptq
Model ArchitectureQwen2VLForConditionalGeneration
Licenseother
Context Length32768
Model Max Length32768
Transformers Version4.45.0.dev0
Tokenizer ClassQwen2Tokenizer
Padding Token<|endoftext|>
Vocabulary Size152064
Torch Data Typefloat16
Errorsreplace

Best Alternatives to Qwen2 VL 72B Instruct GPTQ Int8

Best Alternatives
Context / RAM
Downloads
Likes
...wen2 VL 72B Instruct GPTQ Int432K / 43 GB53630
Qwen2 VL 72B Instruct 8bit32K / 77.8 GB794
Qwen2 VL 72B Instruct 4bit32K / 41.3 GB725
Qwen2 VL 72B Instruct Bnb 4bit32K / 41.8 GB3955
Qwen2 VL 72B Instruct AWQ32K / 43 GB8719050
Qwen2 VL 72B Instruct32K / 146.8 GB60157311
Qwen2 VL 72B Instruct32K / 147.5 GB611
Note: green Score (e.g. "73.2") means that the model is better than Qwen/Qwen2-VL-72B-Instruct-GPTQ-Int8.