LLM EXPLORER 59,358 MODELS INDEXED

Llama 2 7B Chat 4bit Gptq by hoang1123

By hoang1123 · 3 downloads

Llama 2 7B Chat 4bit Gptq is an open-source language model by hoang1123. Features: 7b LLM, VRAM: 3.9GB, Context: 4K, Quantized, LLM Explorer Score: 0.12.

  Arxiv:1910.09700   4-bit   4bit   Endpoints compatible   Gptq   Llama   Quantized   Region:us

Llama 2 7B Chat 4bit Gptq Parameters and Internals

LLM NameLlama 2 7B Chat 4bit Gptq
Repository πŸ€—https://huggingface.co/hoang1123/Llama-2-7b-chat-4bit-gptq 
Base Model(s)  Llama 2 7B Chat Hf 4bits Q   AAProject/Llama-2-7b-chat-hf-4bits-Q
Model Size7b
Required VRAM3.9 GB
Updated2026-07-28
Maintainerhoang1123
Model Typellama
Model Files  3.9 GB
GPTQ QuantizationYes
Quantization Typegptq|4bit
Model ArchitectureLlamaForCausalLM
Context Length4096
Model Max Length4096
Transformers Version4.39.3
Tokenizer ClassLlamaTokenizer
Padding Token<unk>
Vocabulary Size32000
Torch Data Typefloat16

Best Alternatives to Llama 2 7B Chat 4bit Gptq

Best Alternatives
Context / RAM
Downloads
Likes
Yarn Llama 2 7B 128K GPTQ128K / 3.9 GB87
Yarn Llama 2 7B 64K GPTQ64K / 3.9 GB101
... 7B 32K Instructions V4 Marlin32K / 4.1 GB80
Aixcoder 7B GPTQ32K / 4.5 GB41
Calm2 7B Chat GPTQ32K / 4.4 GB65
...Calm2 7B Chat GPTQ Calib Ja 1K32K / 4.4 GB45
Llama 2 7B 32K Instruct GPTQ32K / 3.9 GB4927
Codebear 7B 4bit16K / 3.9 GB41
...a 7B Instruct GPTQ Calib Ja 1K16K / 3.9 GB60
Pandalyst 7B V1.2 GPTQ16K / 3.9 GB81