Llava V1.6 34B 4bit is an open-source language model by mlx-community. Features: 34b LLM, VRAM: 19.5GB, Quantized, LLM Explorer Score: 0.13.
| LLM Name | Llava V1.6 34B 4bit |
| Repository π€ | https://huggingface.co/mlx-community/llava-v1.6-34b-4bit |
| Model Size | 34b |
| Required VRAM | 19.5 GB |
| Updated | 2026-08-08 |
| Maintainer | mlx-community |
| Model Type | llava_next |
| Model Files | |
| Supported Languages | en |
| Quantization Type | 4bit |
| Model Architecture | LlavaNextForConditionalGeneration |
| Model Max Length | 4096 |
| Transformers Version | 4.39.0.dev0 |
| Tokenizer Class | LlamaTokenizer |
| Padding Token | <unk> |
| Vocabulary Size | 64064 |
| Torch Data Type | float16 |
Best Alternatives |
Context / RAM |
Downloads |
Likes |
|---|---|---|---|
| Llava V1.6 34B 8bit | 0K / 36.7 GB | 37 | 2 |