LLM EXPLORER 59,420 MODELS INDEXED

Llama 2 13B 8bit by Peeepy

By Peeepy · 4 downloads

Llama 2 13B 8bit is an open-source language model by Peeepy. Features: 13b LLM, VRAM: 13.4GB, Context: 4K, Quantized, LLM Explorer Score: 0.08.

  8bit   Endpoints compatible   Llama   Quantized   Region:us   Safetensors

Llama 2 13B 8bit Parameters and Internals

LLM NameLlama 2 13B 8bit
Repository πŸ€—https://huggingface.co/Peeepy/llama-2-13b-8bit 
Base Model(s)  ToddLora 13b V2   autobots/ToddLora_13b_v2
Model Size13b
Required VRAM13.4 GB
Updated2026-08-02
MaintainerPeeepy
Model Typellama
Model Files  13.4 GB
Quantization Type8bit
Model ArchitectureLlamaForCausalLM
Context Length4096
Model Max Length4096
Transformers Version4.31.0.dev0
Tokenizer ClassLlamaTokenizer
Beginning of Sentence Token<s>
End of Sentence Token</s>
Unk Token<unk>
Vocabulary Size32000
Torch Data Typefloat16

Best Alternatives to Llama 2 13B 8bit

Best Alternatives
Context / RAM
Downloads
Likes
Llama13b 32K Illumeet Finetune32K / 26 GB50
...Maid V3 13B 32K 8.0bpw H8 EXL232K / 13.2 GB81
...Maid V3 13B 32K 6.0bpw H6 EXL232K / 10 GB51
WhiteRabbitNeo 13B V116K / 26 GB3146453
CodeLlama 13B Python Fp1616K / 26 GB12125
CodeLlama 13B Fp1616K / 26 GB967
CodeLlama 13B Instruct Fp1616K / 26 GB13728
Codellama 13B Bnb 4bit16K / 7.2 GB1295
...Llama 13B Instruct Hf 4bit MLX16K / 7.8 GB1133
WhiteRabbitNeo 13B V1 4bit Mlx16K / 7.8 GB1472
Note: green Score (e.g. "73.2") means that the model is better than Peeepy/llama-2-13b-8bit.