LLM EXPLORER 59,420 MODELS INDEXED

Llama 2 7B Hf 4bit G64 HQQ by mobiuslabsgmbh

By mobiuslabsgmbh · 11 downloads

Llama 2 7B Hf 4bit G64 HQQ is an open-source language model by mobiuslabsgmbh. Features: 7b LLM, VRAM: 4.1GB, Context: 4K, License: llama2, Quantized, LLM Explorer Score: 0.1.

  4bit   Llama   Quantized   Region:us

Llama 2 7B Hf 4bit G64 HQQ Parameters and Internals

Model Type 
text generation
Use Cases 
Limitations:
Only supports single GPU runtime, Not compatible with HuggingFace's PEFT
Input Output 
Accepted Modalities:
text
LLM NameLlama 2 7B Hf 4bit G64 HQQ
Repository πŸ€—https://huggingface.co/mobiuslabsgmbh/Llama-2-7b-hf-4bit_g64-HQQ 
Model Size7b
Required VRAM4.1 GB
Updated2026-08-03
Maintainermobiuslabsgmbh
Model Typellama
Model Files  4.1 GB
Quantization Type4bit
Model ArchitectureLlamaForCausalLM
Licensellama2
Context Length4096
Model Max Length4096
Transformers Version4.34.1
Tokenizer ClassLlamaTokenizer
Padding Token</s>
Vocabulary Size32000
Torch Data Typefloat16

Best Alternatives to Llama 2 7B Hf 4bit G64 HQQ

Best Alternatives
Context / RAM
Downloads
Likes
Smaugv0.1 6.0bpw H6 EXL2195K / 26.4 GB14
Smaugv0.1 5.0bpw H6 EXL2195K / 22.3 GB33
Smaugv0.1 4.65bpw H6 EXL2195K / 20.8 GB51
Smaugv0.1 4.0bpw H6 EXL2195K / 18 GB41
Smaugv0.1 8.0bpw H8 EXL2195K / 34.9 GB41
Smaugv0.1 3.0bpw H6 EXL2195K / 13.9 GB01
DeepSeek Prover V2 7B 4bit64K / 3.9 GB2364
Mistral 7B Openplatypus 1K32K / 29 GB18140
Mistral 7B OpenOrca 1K32K / 29 GB18113
...rnlm2 20B Llama 4.0bpw H6 EXL232K / 11 GB51
Note: green Score (e.g. "73.2") means that the model is better than mobiuslabsgmbh/Llama-2-7b-hf-4bit_g64-HQQ.