LLM EXPLORER 59,420 MODELS INDEXED

NexusRaven V2 13B AWQ by NaiveAttention

By NaiveAttention · 1 downloads

NexusRaven V2 13B AWQ is an open-source language model by NaiveAttention. Features: 13b LLM, VRAM: 7.2GB, Context: 16K, License: apache-2.0, Quantized, LLM Explorer Score: 0.14, Arc: 45.1, HellaSwag: 67.4, MMLU: 44.9, GSM8K: 20.9.

  4-bit   Awq   Dataset:wikitext   Endpoints compatible   Function calling   Llama   Quantized   Region:us   Safetensors

NexusRaven V2 13B AWQ Parameters and Internals

LLM NameNexusRaven V2 13B AWQ
Repository πŸ€—https://huggingface.co/NaiveAttention/NexusRaven-V2-13B-awq 
Model NameNexusRaven V2 13B
Model CreatorNexusflow
Model Size13b
Required VRAM7.2 GB
Updated2026-07-31
MaintainerNaiveAttention
Model Typellama
Model Files  7.2 GB
AWQ QuantizationYes
Quantization Typeawq
Model ArchitectureLlamaForCausalLM
Licenseapache-2.0
Context Length16384
Model Max Length16384
Transformers Version4.37.2
Tokenizer ClassCodeLlamaTokenizer
Padding Token</s>
Vocabulary Size32024
Torch Data Typefloat16

Best Alternatives to NexusRaven V2 13B AWQ

Best Alternatives
Context / RAM
Downloads
Likes
Yarn Llama 2 13B 128K AWQ128K / 7.2 GB42
LongAlign 13B 64K AWQ64K / 7.2 GB42
...oboros L2 13B 2 1 YaRN 64K AWQ64K / 7.2 GB52
OrcaMaid V3 13B 32K AWQ32K / 7.2 GB54
OrcaMaid V2 FIX 13B 32K AWQ32K / 7.2 GB71
NexusRaven V2 13B AWQ16K / 7.2 GB113
...th CodeLlama 13B Python Hf AWQ16K / 7.5 GB60
WhiteRabbitNeo 13B AWQ16K / 7.2 GB3324
NexusRaven V2 13B AWQ16K / 7.2 GB101
Ramgpt 13B AWQ Gemm16K / 7.2 GB01
Note: green Score (e.g. "73.2") means that the model is better than NaiveAttention/NexusRaven-V2-13B-awq.