LLM EXPLORER 59,516 MODELS INDEXED

UndiMix V2 13B AWQ by TheBloke

By TheBloke · 3 downloads

UndiMix V2 13B AWQ is an open-source language model by TheBloke. Features: 13b LLM, VRAM: 7.2GB, Context: 4K, License: cc-by-nc-4.0, Quantized, LLM Explorer Score: 0.09.

  4-bit   Awq Base model:quantized:undi95/un... Base model:undi95/undimix-v2-1...   Fp16   Llama   Quantized   Region:us   Safetensors

UndiMix V2 13B AWQ Parameters and Internals

Model Type 
llama
Additional Notes 
The model supports AWQ quantization and can be used efficiently with smaller GPUs. Multiple quantization options are available for different inference settings.
Input Output 
Input Format:
Below is an instruction that describes a task. Write a response that appropriately completes the request. ### Instruction: {prompt} ### Response:
LLM NameUndiMix V2 13B AWQ
Repository πŸ€—https://huggingface.co/TheBloke/UndiMix-v2-13B-AWQ 
Model NameUndiMix v2 13B
Model CreatorUndi95
Base Model(s)  UndiMix V2 13B   Undi95/UndiMix-v2-13b
Model Size13b
Required VRAM7.2 GB
Updated2026-07-18
MaintainerTheBloke
Model Typellama
Model Files  7.2 GB
AWQ QuantizationYes
Quantization Typefp16|awq
Model ArchitectureLlamaForCausalLM
Licensecc-by-nc-4.0
Context Length4096
Model Max Length4096
Transformers Version4.32.1
Tokenizer ClassLlamaTokenizer
Beginning of Sentence Token<s>
End of Sentence Token</s>
Unk Token<unk>
Vocabulary Size32000
Torch Data Typefloat16

Best Alternatives to UndiMix V2 13B AWQ

Best Alternatives
Context / RAM
Downloads
Likes
Yarn Llama 2 13B 128K AWQ128K / 7.2 GB42
LongAlign 13B 64K AWQ64K / 7.2 GB42
...oboros L2 13B 2 1 YaRN 64K AWQ64K / 7.2 GB52
OrcaMaid V3 13B 32K AWQ32K / 7.2 GB54
OrcaMaid V2 FIX 13B 32K AWQ32K / 7.2 GB71
NexusRaven V2 13B AWQ16K / 7.2 GB113
NexusRaven V2 13B AWQ16K / 7.2 GB13
...th CodeLlama 13B Python Hf AWQ16K / 7.5 GB60
WhiteRabbitNeo 13B AWQ16K / 7.2 GB3324
NexusRaven V2 13B AWQ16K / 7.2 GB101
Note: green Score (e.g. "73.2") means that the model is better than TheBloke/UndiMix-v2-13B-AWQ.