LLM EXPLORER 60,384 MODELS INDEXED

GLM 5.3 Flash AWQ W4A16 by wtdcode

By wtdcode · 4247 downloads

GLM 5.3 Flash AWQ W4A16 is an open-source language model by wtdcode. Features: 321.3b LLM, VRAM: 176GB, Quantized, LLM Explorer Score: 0.51, ELO: 1474.

  Awq Base model:quantized:zai-org/g... Base model:zai-org/glm-5.3-fla...   Compressed-tensors   Glm5 next   Quantized   Region:us   Safetensors   Sharded   Tensorflow

GLM 5.3 Flash AWQ W4A16 Benchmarks

nn.n% — How the model compares to the reference models: Anthropic Sonnet 3.5 ("so35"), GPT-4o ("gpt4o") or GPT-4 ("gpt4").

GLM 5.3 Flash AWQ W4A16 Parameters and Internals

LLM NameGLM 5.3 Flash AWQ W4A16
Repository πŸ€—https://huggingface.co/wtdcode/GLM-5.3-Flash-AWQ-W4A16 
Base Model(s)  zai-org/GLM-5.3-Flash   zai-org/GLM-5.3-Flash
Model Size321.3b
Required VRAM176 GB
Updated2026-09-04
Maintainerwtdcode
Model Typeglm5_next
Model Files  20.0 GB: 1-of-9   20.0 GB: 2-of-9   20.0 GB: 3-of-9   20.0 GB: 4-of-9   20.0 GB: 5-of-9   20.0 GB: 6-of-9   20.0 GB: 7-of-9   20.0 GB: 8-of-9   16.0 GB: 9-of-9   4.3 GB   4.3 GB   4.3 GB   1.9 GB
AWQ QuantizationYes
Quantization Typeawq
Model ArchitectureGlm5NextForConditionalGeneration
Model Max Length1048576
Transformers Version5.16.1
Tokenizer ClassTokenizersBackend
Padding Token<|endoftext|>

Best Alternatives to GLM 5.3 Flash AWQ W4A16

Best Alternatives
Context / RAM
Downloads
Likes
GLM 5.3 Flash JANG MTP0K / 102.8 GB2771
...5.3 Flash Abliterated MLX 4bit0K / 185.6 GB9112
GLM 5.3 Flash Uncensored AWQ0K / 176 GB641
GLM 5.3 Flash0K / 237.5 GB341289
GLM 5.3 Flash BF160K / 237.1 GB2438
GLM 5.3 Flash FP80K / 237.5 GB504822
GLM 5.3 Flash Nota NVFP40K / 21 GB16017
GLM 5.3 Flash0K / 242.5 GB103415
...Flash MXFP4 Mixed CT AutoRound0K / 67.4 GB1951
GLM 5.3 Flash DERISKED BF160K / 237.1 GB1343
Note: green Score (e.g. "73.2") means that the model is better than wtdcode/GLM-5.3-Flash-AWQ-W4A16.