LLM EXPLORER 59,420 MODELS INDEXED

LLaMA 13B 4bit 32g by Neko-Institute-of-Science

By Neko-Institute-of-Science · 7 downloads

LLaMA 13B 4bit 32g is an open-source language model by Neko-Institute-of-Science. Features: 13b LLM, VRAM: 8GB, Context: 2K, Quantized, LLM Explorer Score: 0.06.

  4bit   Endpoints compatible   Llama   Quantized   Region:us

LLaMA 13B 4bit 32g Parameters and Internals

Additional Notes 
The performance metrics are presented for multiple configurations of the model using different datasets as benchmarks.
LLM NameLLaMA 13B 4bit 32g
Repository πŸ€—https://huggingface.co/Neko-Institute-of-Science/LLaMA-13B-4bit-32g 
Base Model(s)  ToddLora 13b V2   autobots/ToddLora_13b_v2
Model Size13b
Required VRAM8 GB
Updated2026-08-04
MaintainerNeko-Institute-of-Science
Model Typellama
Model Files  8.0 GB
Quantization Type4bit
Model ArchitectureLlamaForCausalLM
Context Length2048
Model Max Length2048
Transformers Version4.28.1
Tokenizer ClassLlamaTokenizer
Beginning of Sentence Token<s>
End of Sentence Token</s>
Unk Token<unk>
Vocabulary Size32000
Torch Data Typefloat16

Best Alternatives to LLaMA 13B 4bit 32g

Best Alternatives
Context / RAM
Downloads
Likes
Llama13b 32K Illumeet Finetune32K / 26 GB50
...Maid V3 13B 32K 8.0bpw H8 EXL232K / 13.2 GB81
...Maid V3 13B 32K 6.0bpw H6 EXL232K / 10 GB51
WhiteRabbitNeo 13B V116K / 26 GB3146453
CodeLlama 13B Python Fp1616K / 26 GB12125
CodeLlama 13B Fp1616K / 26 GB967
CodeLlama 13B Instruct Fp1616K / 26 GB13728
Codellama 13B Bnb 4bit16K / 7.2 GB1295
...Llama 13B Instruct Hf 4bit MLX16K / 7.8 GB1133
WhiteRabbitNeo 13B V1 4bit Mlx16K / 7.8 GB1472
Note: green Score (e.g. "73.2") means that the model is better than Neko-Institute-of-Science/LLaMA-13B-4bit-32g.