LLM EXPLORER 60,645 MODELS INDEXED

Qwen2.5 VL 7B Instruct FP4 by nvidia

By nvidia · 49165 downloads

Qwen2.5 VL 7B Instruct FP4 is an open-source language model by nvidia. Features: 7b LLM, VRAM: 7.2GB, Context: 125K, License: other, Instruction-Based, LLM Explorer Score: 0.26.

  8-bit Base model:quantized:qwen/qwen... Base model:qwen/qwen2.5-vl-7b-...   Conversational   Fp4   Instruct   Model optimizer   Modelopt   Nvidia   Quantized   Qwen2 5 vl   Region:us   Safetensors   Sharded   Tensorflow

Qwen2.5 VL 7B Instruct FP4 Parameters and Internals

LLM NameQwen2.5 VL 7B Instruct FP4
Repository πŸ€—https://huggingface.co/nvidia/Qwen2.5-VL-7B-Instruct-FP4 
Base Model(s)  Qwen/Qwen2.5-VL-7B-Instruct   Qwen/Qwen2.5-VL-7B-Instruct
Model Size7b
Required VRAM7.2 GB
Updated2025-10-10
Maintainernvidia
Model Typeqwen2_5_vl
Instruction-BasedYes
Model Files  5.0 GB: 1-of-2   2.2 GB: 2-of-2
Model ArchitectureQwen2_5_VLForConditionalGeneration
Licenseother
Context Length128000
Model Max Length128000
Transformers Version4.56.1
Tokenizer ClassQwen2Tokenizer
Padding Token<|endoftext|>
Vocabulary Size152064
Torch Data Typebfloat16
Errorsreplace

Best Alternatives to Qwen2.5 VL 7B Instruct FP4

Best Alternatives
Context / RAM
Downloads
Likes
OmniLong Qwen2.5 VL 7B512K / 16.6 GB72
Qwen2.5 VL 7B Instruct125K / 16.7 GB87409051666
OlmOCR 2 7B 1025 FP8125K / 10 GB540426253
OlmOCR 2 7B 1025125K / 16.6 GB244728157
RolmOCR125K / 16.6 GB224280589
Cosmos Reason1 7B125K / 16.6 GB62913244
SpaSEViLA Cold Start 7B125K / 16.6 GB190
Zoom IQA 7B125K / 16.6 GB141
OlmOCR 7B 0825 FP8125K / 10 GB3801110
Embodied Navigator 7B GRPO125K / 17.3 GB201
Note: green Score (e.g. "73.2") means that the model is better than nvidia/Qwen2.5-VL-7B-Instruct-FP4.