LLM EXPLORER 59,516 MODELS INDEXED

Orca2myth7.2 AWQ by TheBloke

By TheBloke · 7 downloads

Orca2myth7.2 AWQ is an open-source language model by TheBloke. Features: 20b LLM, VRAM: 10.9GB, Context: 4K, License: other, Quantized, LLM Explorer Score: 0.1.

  4-bit   Awq Base model:quantized:thebigble... Base model:thebigblender/orca2...   Llama   Quantized   Region:us   Safetensors   Sharded   Tensorflow

Orca2myth7.2 AWQ Parameters and Internals

Model Type 
llama
Additional Notes 
The model was quantized using hardware provided by Massed Compute. AWQ supports efficient, low-bit weight quantization enhancing Transformer-based inference speed and quality.
Training Details 
Data Sources:
VMware Open Instruct
Methodology:
Float16 quantization, AWQ (4-bit quantization)
Context Length:
4096
Hardware Used:
Massed Compute
Model Architecture:
A merge of Orca2flat and PygmalionAI/mythalion-13b using specific layer ranges and merge methods.
Input Output 
Input Format:
{prompt}
LLM NameOrca2myth7.2 AWQ
Repository πŸ€—https://huggingface.co/TheBloke/Orca2myth7.2-AWQ 
Model NameOrca2Myth7.2
Model CreatorThe Big Blender
Base Model(s)  TheBigBlender/Orca2myth7.2   TheBigBlender/Orca2myth7.2
Model Size20b
Required VRAM10.9 GB
Updated2026-08-07
MaintainerTheBloke
Model Typellama
Model Files  10.0 GB: 1-of-2   0.9 GB: 2-of-2
AWQ QuantizationYes
Quantization Typeawq
Model ArchitectureLlamaForCausalLM
Licenseother
Context Length4096
Model Max Length4096
Transformers Version4.35.2
Tokenizer ClassLlamaTokenizer
Vocabulary Size32000
Torch Data Typefloat16

Best Alternatives to Orca2myth7.2 AWQ

Best Alternatives
Context / RAM
Downloads
Likes
Norocetacean 20B 10K AWQ10K / 10.9 GB112
DaringMaid 20B AWQ4K / 10.9 GB22
Rose Kimiko 20B AWQ4K / 10.9 GB01
Nethena 20B Glued AWQ4K / 10.9 GB42
Iambe Storyteller 20B AWQ4K / 10.9 GB71
Iambe 20B DARE AWQ4K / 10.9 GB31
Noromaid 20B V0.1.1 AWQ4K / 10.9 GB54
Rose 20B AWQ4K / 10.9 GB11
Nethena 20B AWQ4K / 10.9 GB82
MXLewd L2 20B AWQ4K / 10.9 GB93