LLM EXPLORER 59,358 MODELS INDEXED

Llama 2 7B Chat Hf 4bits Q by AAProject

By AAProject · 8 downloads

Llama 2 7B Chat Hf 4bits Q is an open-source language model by AAProject. Features: 7b LLM, VRAM: 4.2GB, Context: 4K, LLM Explorer Score: 0.12.

  Arxiv:1910.09700   4-bit   Bitsandbytes   Conversational   Endpoints compatible   Llama   Region:us   Safetensors

Llama 2 7B Chat Hf 4bits Q Parameters and Internals

LLM NameLlama 2 7B Chat Hf 4bits Q
Repository πŸ€—https://huggingface.co/AAProject/Llama-2-7b-chat-hf-4bits-Q 
Model Size7b
Required VRAM4.2 GB
Updated2026-08-24
MaintainerAAProject
Model Typellama
Model Files  4.2 GB
Model ArchitectureLlamaForCausalLM
Context Length4096
Model Max Length4096
Transformers Version4.42.0.dev0
Tokenizer ClassLlamaTokenizer
Vocabulary Size32000
Torch Data Typefloat16

Quantized Models of the Llama 2 7B Chat Hf 4bits Q

Model
Likes
Downloads
VRAM
Llama 2 7B Chat 4bit Gptq133 GB
...lama 2 7B Chat Hf 4bit G64 HQQ3124 GB

Best Alternatives to Llama 2 7B Chat Hf 4bits Q

Best Alternatives
Context / RAM
Downloads
Likes
1241024K / 16.1 GB930
1621024K / 16.1 GB600
1571024K / 16.1 GB1010
1181024K / 16.1 GB150
A5.41024K / 16.1 GB120
A3.41024K / 16.1 GB130
A2.41024K / 16.1 GB120
A6 L1024K / 16.1 GB2010
M1024K / 16.1 GB1270
2 Very Sci Fi1024K / 16.1 GB3170
Note: green Score (e.g. "73.2") means that the model is better than AAProject/Llama-2-7b-chat-hf-4bits-Q.