LLM EXPLORER 59,420 MODELS INDEXED

Qwen3 1.7B MATH GDPO by wzx111

By wzx111 · 23 downloads

Qwen3 1.7B MATH GDPO is an open-source language model by wzx111. Features: 1.7b LLM, VRAM: 3.4GB, Context: 40K, LLM Explorer Score: 0.18.

  Arxiv:2402.03300 Base model:finetune:qwen/qwen3...   Base model:qwen/qwen3-1.7b   Conversational Dataset:watermelonhjg/math-lig...   Endpoints compatible   Gdpo   Generated from trainer   Open-r1   Qwen3   Region:us   Safetensors   Trl

Qwen3 1.7B MATH GDPO Parameters and Internals

LLM NameQwen3 1.7B MATH GDPO
Repository πŸ€—https://huggingface.co/wzx111/Qwen3-1.7B-MATH-GDPO 
Model NameQwen3-1.7B-MATH-GDPO
Base Model(s)  Qwen/Qwen3-1.7B   Qwen/Qwen3-1.7B
Model Size1.7b
Required VRAM3.4 GB
Updated2026-06-04
Maintainerwzx111
Model Typeqwen3
Model Files  3.4 GB   0.0 GB
Model ArchitectureQwen3ForCausalLM
Context Length40960
Model Max Length40960
Transformers Version4.52.0.dev0
Tokenizer ClassQwen2Tokenizer
Padding Token<|endoftext|>
Vocabulary Size151936
Torch Data Typebfloat16
Errorsreplace

Best Alternatives to Qwen3 1.7B MATH GDPO

Best Alternatives
Context / RAM
Downloads
Likes
Lucy 128K128K / 3.4 GB123112
Polaris 1.7B Preview128K / 3.4 GB738
DictaLM 3.0 1.7B Instruct60K / 3.4 GB28421
Qwen3 1.7B40K / 4 GB5633291500
DualMind TKD Agentic 1.7B40K / 3.4 GB17980
Atomight V2.5 1.7B40K / 3.4 GB14061
Qwen3 1.7B Base MED ChatVector40K / 3.4 GB2041
... 1.7B Base MED ChatVector 081240K / 3.4 GB491
Qwen3 1.7B Base MED ChatVector40K / 3.4 GB461
Qwen3 1.7B Base MED ChatVector40K / 3.4 GB441
Note: green Score (e.g. "73.2") means that the model is better than wzx111/Qwen3-1.7B-MATH-GDPO.