LLM EXPLORER 58,178 MODELS INDEXED

Qwen3 1.7B Grpo Training 1st Half Epoch by shenzhentianyi

By shenzhentianyi · 6 downloads

Qwen3 1.7B Grpo Training 1st Half Epoch is an open-source language model by shenzhentianyi. Features: 1.7b LLM, VRAM: 3.4GB, Context: 32K, LLM Explorer Score: 0.19.

  Qwen3   Region:us   Safetensors

Qwen3 1.7B Grpo Training 1st Half Epoch Parameters and Internals

LLM NameQwen3 1.7b Grpo Training 1st Half Epoch
Repository πŸ€—https://huggingface.co/shenzhentianyi/qwen3_1.7b_grpo_training_1st_half_epoch 
Model Size1.7b
Required VRAM3.4 GB
Updated2025-09-18
Maintainershenzhentianyi
Model Typeqwen3
Model Files  3.4 GB
Model ArchitectureQwen3ForCausalLM
Context Length32768
Model Max Length32768
Transformers Version4.53.2
Tokenizer ClassQwen2Tokenizer
Padding Token<|vision_pad|>
Vocabulary Size151936
Torch Data Typefloat16
Errorsreplace

Best Alternatives to Qwen3 1.7B Grpo Training 1st Half Epoch

Best Alternatives
Context / RAM
Downloads
Likes
Lucy 128K128K / 3.4 GB123112
Polaris 1.7B Preview128K / 3.4 GB738
DictaLM 3.0 1.7B Instruct60K / 3.4 GB28421
Qwen3 1.7B40K / 4 GB5633291500
DualMind TKD Agentic 1.7B40K / 3.4 GB17980
Atomight V2.5 1.7B40K / 3.4 GB14061
MedPsy 1.7B40K / 4.1 GB9338
Supertron2 1.7B40K / 3.4 GB2487
Qwen3 1.7B40K / 3.4 GB12293812
Safety Model40K / 3.4 GB2310
Note: green Score (e.g. "73.2") means that the model is better than shenzhentianyi/qwen3_1.7b_grpo_training_1st_half_epoch.