LLM EXPLORER 59,420 MODELS INDEXED

OrcaMaid V2 FIX 13B 32K AWQ by TheBloke

By TheBloke · 7 downloads

OrcaMaid V2 FIX 13B 32K AWQ is an open-source language model by TheBloke. Features: 13b LLM, VRAM: 7.2GB, Context: 32K, License: other, Quantized, LLM Explorer Score: 0.1.

  4-bit   Awq Base model:ddh0/orcamaid-v2-fi... Base model:quantized:ddh0/orca...   Custom code   Llama   Quantized   Region:us   Safetensors

OrcaMaid V2 FIX 13B 32K AWQ Parameters and Internals

Model Type 
llama, text-generation
Additional Notes 
Extended context length to 32K via YaRN. Issues in the previous tokenizer version were rectified.
Input Output 
Input Format:
Below is an instruction that describes a task. Write a response that appropriately completes the request. ### Instruction: {prompt} ### Response:
LLM NameOrcaMaid V2 FIX 13B 32K AWQ
Repository πŸ€—https://huggingface.co/TheBloke/OrcaMaid-v2-FIX-13B-32k-AWQ 
Model NameOrcamaid V2 Fix 13B 32K
Model Creatorddh0
Base Model(s)  OrcaMaid V2 FIX 13B 32K   ddh0/OrcaMaid-v2-FIX-13b-32k
Model Size13b
Required VRAM7.2 GB
Updated2026-07-06
MaintainerTheBloke
Model Typellama
Model Files  7.2 GB
AWQ QuantizationYes
Quantization Typeawq
Model ArchitectureLlamaForCausalLM
Licenseother
Context Length32768
Model Max Length32768
Transformers Version4.35.2
Tokenizer ClassLlamaTokenizer
Vocabulary Size32000
Torch Data Typefloat16

Best Alternatives to OrcaMaid V2 FIX 13B 32K AWQ

Best Alternatives
Context / RAM
Downloads
Likes
Yarn Llama 2 13B 128K AWQ128K / 7.2 GB42
LongAlign 13B 64K AWQ64K / 7.2 GB42
...oboros L2 13B 2 1 YaRN 64K AWQ64K / 7.2 GB52
OrcaMaid V3 13B 32K AWQ32K / 7.2 GB54
NexusRaven V2 13B AWQ16K / 7.2 GB113
NexusRaven V2 13B AWQ16K / 7.2 GB13
...th CodeLlama 13B Python Hf AWQ16K / 7.5 GB60
WhiteRabbitNeo 13B AWQ16K / 7.2 GB3324
NexusRaven V2 13B AWQ16K / 7.2 GB101
Ramgpt 13B AWQ Gemm16K / 7.2 GB01
Note: green Score (e.g. "73.2") means that the model is better than TheBloke/OrcaMaid-v2-FIX-13B-32k-AWQ.