LLM EXPLORER 60,220 MODELS INDEXED

GLM 5.3 Flash W4A16 AutoRound by Intel

By Intel · 104 downloads

GLM 5.3 Flash W4A16 AutoRound is an open-source language model by Intel. Features: 50.3b LLM, VRAM: 178.4GB, License: mit, LLM Explorer Score: 0.34.

  Arxiv:2309.05516   4-bit   Auto-round Base model:quantized:zai-org/g... Base model:zai-org/glm-5.3-fla...   Conversational   En   Endpoints compatible   Glm5 next   Image-text-to-text   Region:us   Safetensors   Sharded   Tensorflow   Zh

GLM 5.3 Flash W4A16 AutoRound Parameters and Internals

LLM NameGLM 5.3 Flash W4A16 AutoRound
Repository πŸ€—https://huggingface.co/Intel/GLM-5.3-Flash-W4A16-AutoRound 
Base Model(s)  zai-org/GLM-5.3-Flash   zai-org/GLM-5.3-Flash
Model Size50.3b
Required VRAM178.4 GB
Updated2026-09-02
MaintainerIntel
Model Typeglm5_next
Model Files  5.4 GB: 1-of-34   5.4 GB: 2-of-34   5.4 GB: 3-of-34   5.4 GB: 4-of-34   5.4 GB: 5-of-34   5.4 GB: 6-of-34   5.4 GB: 7-of-34   5.4 GB: 8-of-34   5.4 GB: 9-of-34   5.4 GB: 10-of-34   5.4 GB: 11-of-34   5.4 GB: 12-of-34   5.4 GB: 13-of-34   5.4 GB: 14-of-34   5.4 GB: 15-of-34   5.4 GB: 16-of-34   5.4 GB: 17-of-34   5.4 GB: 18-of-34   5.4 GB: 19-of-34   5.4 GB: 20-of-34   5.4 GB: 21-of-34   5.4 GB: 22-of-34   5.4 GB: 23-of-34   5.4 GB: 24-of-34   5.4 GB: 25-of-34   5.4 GB: 26-of-34   5.4 GB: 27-of-34   5.4 GB: 28-of-34   5.4 GB: 29-of-34   5.4 GB: 30-of-34   5.4 GB: 31-of-34   5.4 GB: 32-of-34   4.3 GB: 33-of-34   1.3 GB: 34-of-34   0.0 GB   4.1 GB
Supported Languagesen zh
Model ArchitectureGlm5NextForConditionalGeneration
Licensemit
Model Max Length1048576
Transformers Version5.16.0
Tokenizer ClassTokenizersBackend
Padding Token<|endoftext|>