LLM EXPLORER 59,516 MODELS INDEXED

MXLewd L2 20B AWQ by TheBloke

By TheBloke · 9 downloads

MXLewd L2 20B AWQ is an open-source language model by TheBloke. Features: 20b LLM, VRAM: 10.9GB, Context: 4K, License: cc-by-nc-4.0, Quantized, LLM Explorer Score: 0.09.

  4-bit   Awq Base model:quantized:undi95/mx... Base model:undi95/mxlewd-l2-20...   Llama   Quantized   Region:us   Safetensors   Sharded   Tensorflow

MXLewd L2 20B AWQ Parameters and Internals

Model Type 
llama
Additional Notes 
The MXLewd L2 20B model has been quantized using the AWQ method for efficient and accurate low-bit weight quantization, specifically supporting 4-bit quantization. This allows faster transformer-based inference and the use of smaller GPUs for deployment and cost savings.
LLM NameMXLewd L2 20B AWQ
Repository πŸ€—https://huggingface.co/TheBloke/MXLewd-L2-20B-AWQ 
Model NameMXLewd L2 20B
Model CreatorUndi
Base Model(s)  MXLewd L2 20B   Undi95/MXLewd-L2-20B
Model Size20b
Required VRAM10.9 GB
Updated2026-05-19
MaintainerTheBloke
Model Typellama
Model Files  10.0 GB: 1-of-2   0.9 GB: 2-of-2
AWQ QuantizationYes
Quantization Typeawq
Model ArchitectureLlamaForCausalLM
Licensecc-by-nc-4.0
Context Length4096
Model Max Length4096
Transformers Version4.33.2
Tokenizer ClassLlamaTokenizer
Beginning of Sentence Token<s>
End of Sentence Token</s>
Unk Token<unk>
Vocabulary Size32000
Torch Data Typebfloat16

Best Alternatives to MXLewd L2 20B AWQ

Best Alternatives
Context / RAM
Downloads
Likes
Norocetacean 20B 10K AWQ10K / 10.9 GB112
Orca2myth7.2 AWQ4K / 10.9 GB73
DaringMaid 20B AWQ4K / 10.9 GB22
Rose Kimiko 20B AWQ4K / 10.9 GB01
Nethena 20B Glued AWQ4K / 10.9 GB42
Iambe Storyteller 20B AWQ4K / 10.9 GB71
Iambe 20B DARE AWQ4K / 10.9 GB31
Noromaid 20B V0.1.1 AWQ4K / 10.9 GB54
Rose 20B AWQ4K / 10.9 GB11
Nethena 20B AWQ4K / 10.9 GB82
Note: green Score (e.g. "73.2") means that the model is better than TheBloke/MXLewd-L2-20B-AWQ.