LLM EXPLORER 62,155 MODELS INDEXED

Llama 7B 4bit by kuleshov

By kuleshov · 68 downloads

Llama 7B 4bit is an open-source language model by kuleshov. Features: 7b LLM, VRAM: 3.8GB, Context: 2K, Quantized, LLM Explorer Score: 0.06.

  4bit   Endpoints compatible   Llama   Pytorch   Quantized   Region:us
Model Card on HF πŸ€—: https://huggingface.co/kuleshov/llama-7b-4bit 

Llama 7B 4bit Parameters and Internals

LLM NameLlama 7B 4bit
Repository πŸ€—https://huggingface.co/kuleshov/llama-7b-4bit 
Base Model(s)  Todd Proxy LoRA 7b   autobots/Todd_Proxy_LoRA_7b
Model Size7b
Required VRAM3.8 GB
Updated2026-08-08
Maintainerkuleshov
Model Typellama
Model Files  3.8 GB   4.0 GB
Quantization Type4bit
Model ArchitectureLlamaForCausalLM
Context Length2048
Model Max Length2048
Transformers Version4.28.0
Vocabulary Size32000
Torch Data Typefloat16

Best Alternatives to Llama 7B 4bit

Best Alternatives
Context / RAM
Downloads
Likes
Smaugv0.1 6.0bpw H6 EXL2195K / 26.4 GB14
Smaugv0.1 5.0bpw H6 EXL2195K / 22.3 GB33
Smaugv0.1 4.65bpw H6 EXL2195K / 20.8 GB51
Smaugv0.1 4.0bpw H6 EXL2195K / 18 GB41
Smaugv0.1 8.0bpw H8 EXL2195K / 34.9 GB41
Smaugv0.1 3.0bpw H6 EXL2195K / 13.9 GB01
DeepSeek Prover V2 7B 4bit64K / 3.9 GB2364
Mistral 7B Openplatypus 1K32K / 29 GB18140
Mistral 7B OpenOrca 1K32K / 29 GB18113
...rnlm2 20B Llama 4.0bpw H6 EXL232K / 11 GB51
Note: green Score (e.g. "73.2") means that the model is better than kuleshov/llama-7b-4bit.