Holo1 7B Bf16 is an open-source language model by mlx-community. Features: 7b LLM, VRAM: 16.6GB, Context: 125K, License: apache-2.0, Instruction-Based.
| LLM Name | Holo1 7B Bf16 |
| Repository π€ | https://huggingface.co/mlx-community/Holo1-7B-bf16 |
| Base Model(s) | |
| Model Size | 7b |
| Required VRAM | 16.6 GB |
| Updated | 2026-08-08 |
| Maintainer | mlx-community |
| Model Type | qwen2_5_vl |
| Instruction-Based | Yes |
| Model Files | |
| Supported Languages | en |
| Model Architecture | Qwen2_5_VLForConditionalGeneration |
| License | apache-2.0 |
| Context Length | 128000 |
| Model Max Length | 128000 |
| Transformers Version | 4.51.3 |
| Padding Token | <|endoftext|> |
| Vocabulary Size | 152064 |
| Errors | replace |
Best Alternatives |
Context / RAM |
Downloads |
Likes |
|---|---|---|---|
| OmniLong Qwen2.5 VL 7B | 512K / 16.6 GB | 7 | 2 |
| Qwen2.5 VL 7B Instruct FP4 | 125K / 7.2 GB | 49165 | 3 |
| Qwen2.5 VL 7B Instruct NVFP4 | 125K / 7.2 GB | 8092 | 16 |
| Qwen2.5 VL 7B Instruct FP8 | 125K / 10 GB | 1009 | 8 |
| HuatuoGPT Vision 7B Qwen2.5VL | 125K / 16.6 GB | 358 | 11 |
| OlmOCR 2 7B 1025 Bf16 | 125K / 16.6 GB | 154 | 2 |
| Qwen2.5 VL 7B Instruct Bf16 | 125K / 16.6 GB | 160 | 3 |
| ZwZ 7B | 125K / 16.6 GB | 32 | 10 |
| ViLaSR | 125K / 16.7 GB | 38 | 18 |
| GUI Owl 7B | 32K / 16.7 GB | 698 | 53 |