LLM EXPLORER 60,220 MODELS INDEXED

GLM 5.3 Flash W4A16 NVFP4 K32 Experts FP8 WO by ormandj

By ormandj · 28 downloads

GLM 5.3 Flash W4A16 NVFP4 K32 Experts FP8 WO is an open-source language model by ormandj. Features: 165.5b LLM, VRAM: 83.2GB, License: mit, LLM Explorer Score: 0.32.

  8-bit Base model:quantized:zai-org/g... Base model:zai-org/glm-5.3-fla...   Conversational   Endpoints compatible   Fp8   Glm5 next   Image-text-to-text   Modelopt   Modelopt mixed   Nvfp4   Region:us   Safetensors   Sglang   Sharded   Tensorflow

Best Alternatives to GLM 5.3 Flash W4A16 NVFP4 K32 Experts FP8 WO

Best Alternatives
Context / RAM
Downloads
Likes
GLM 5.3 Flash NVFP40K / 73.1 GB035
GLM 5.3 Flash NVFP40K / 137.3 GB511
GLM 5.3 Flash UNCENSORED NVFP40K / 69.8 GB321027
Note: green Score (e.g. "73.2") means that the model is better than ormandj/GLM-5.3-Flash-W4A16-NVFP4-K32-Experts-FP8-WO.