LLM EXPLORER 59,420 MODELS INDEXED

OrcaMaid V3 13B 32K 8.0bpw H8 EXL2 by LoneStriker

By LoneStriker · 8 downloads

OrcaMaid V3 13B 32K 8.0bpw H8 EXL2 is an open-source language model by LoneStriker. Features: 13b LLM, VRAM: 13.2GB, Context: 32K, License: other, Quantized, Merged, LLM Explorer Score: 0.1.

  Merged Model   Custom code   Endpoints compatible   Exl2   Llama   Quantized   Region:us   Safetensors   Sharded   Tensorflow

OrcaMaid V3 13B 32K 8.0bpw H8 EXL2 Parameters and Internals

Model Type 
text generation
Additional Notes 
This is the third version of OrcaMaid, a weighted gradient SLERP merge aimed at creating an intelligent and human-like model, especially for role-playing (RP).
Input Output 
Input Format:
Alpaca prompt format
Performance Tips:
Customize the system prompt to your specific needs for best results.
LLM NameOrcaMaid V3 13B 32K 8.0bpw H8 EXL2
Repository πŸ€—https://huggingface.co/LoneStriker/OrcaMaid-v3-13b-32k-8.0bpw-h8-exl2 
Merged ModelYes
Model Size13b
Required VRAM13.2 GB
Updated2026-08-08
MaintainerLoneStriker
Model Typellama
Model Files  8.6 GB: 1-of-2   4.6 GB: 2-of-2
Quantization Typeexl2
Model ArchitectureLlamaForCausalLM
Licenseother
Context Length32768
Model Max Length32768
Transformers Version4.36.2
Tokenizer ClassLlamaTokenizer
Padding Token[PAD]
Vocabulary Size32000
Torch Data Typefloat16

Best Alternatives to OrcaMaid V3 13B 32K 8.0bpw H8 EXL2

Best Alternatives
Context / RAM
Downloads
Likes
Llama13b 32K Illumeet Finetune32K / 26 GB50
...Maid V3 13B 32K 6.0bpw H6 EXL232K / 10 GB51
WhiteRabbitNeo 13B V116K / 26 GB3146453
CodeLlama 13B Python Fp1616K / 26 GB12125
CodeLlama 13B Fp1616K / 26 GB967
CodeLlama 13B Instruct Fp1616K / 26 GB13728
Codellama 13B Bnb 4bit16K / 7.2 GB1295
...Llama 13B Instruct Hf 4bit MLX16K / 7.8 GB1133
WhiteRabbitNeo 13B V1 4bit Mlx16K / 7.8 GB1472
Trinity 13B16K / 26 GB1415
Note: green Score (e.g. "73.2") means that the model is better than LoneStriker/OrcaMaid-v3-13b-32k-8.0bpw-h8-exl2.