LLM EXPLORER 56,604 MODELS INDEXED

LongReward Glm4 9B DPO by zai-org

By zai-org · 107 downloads

LongReward Glm4 9B DPO is an open-source language model by zai-org. Features: 9b LLM, VRAM: 18.8GB, Context: 64K, License: other, LLM Explorer Score: 0.15.

  Arxiv:2410.21252 Base model:finetune:zai-org/gl... Base model:zai-org/glm-4-9b-ch...   Chatglm   Conversational   Dataset:thudm/longreward-10k   En   Glm   Region:us   Safetensors   Sharded   Tensorflow   Zh

LongReward Glm4 9B DPO Parameters and Internals

LLM NameLongReward Glm4 9B DPO
Repository πŸ€—https://huggingface.co/zai-org/LongReward-glm4-9b-DPO 
Base Model(s)  Glm 4 9B Chat Hf   THUDM/glm-4-9b-chat-hf
Model Size9b
Required VRAM18.8 GB
Updated2026-08-08
Maintainerzai-org
Model Typeglm
Model Files  5.0 GB: 1-of-4   4.9 GB: 2-of-4   4.9 GB: 3-of-4   4.0 GB: 4-of-4
Supported Languagesen zh
Model ArchitectureGlmForCausalLM
Licenseother
Context Length65536
Model Max Length65536
Transformers Version4.46.0.dev0
Tokenizer ClassPreTrainedTokenizerFast
Padding Token<|endoftext|>
Vocabulary Size151552
Torch Data Typebfloat16

Best Alternatives to LongReward Glm4 9B DPO

Best Alternatives
Context / RAM
Downloads
Likes
Glm 4 9B Chat 1M Hf1024K / 19 GB111412
Glm 4 9B Chat 1M Hf1024K / 19 GB186214
Glm 4 9B Chat Hf128K / 18.8 GB508514
Glm 4 9B Chat Hf128K / 18.8 GB1779025
Glm 4 9B Hf8K / 18.8 GB158810
Glm 4 9B Hf8K / 18.8 GB8537
Note: green Score (e.g. "73.2") means that the model is better than zai-org/LongReward-glm4-9b-DPO.