LLM EXPLORER 62,076 MODELS INDEXED

GRPO 4 70 by BRlkl

By BRlkl · 13 downloads

GRPO 4 70 is an open-source language model by BRlkl. Features: 4b LLM, VRAM: 8.1GB, Context: 256K, License: apache-2.0, LLM Explorer Score: 0.36, ELO: 1411.

Base model:brlkl/orchestrator-... Base model:finetune:brlkl/orch...   Conversational   En   Endpoints compatible   Grpo   Qwen3   Region:us   Safetensors   Sharded   Tensorflow   Trl   Unsloth
Model Card on HF πŸ€—: https://huggingface.co/BRlkl/GRPO-4_70 

GRPO 4 70 Benchmarks

nn.n% — How the model compares to the reference models: Anthropic Sonnet 3.5 ("so35"), GPT-4o ("gpt4o") or GPT-4 ("gpt4").

GRPO 4 70 Parameters and Internals

LLM NameGRPO 4 70
Repository πŸ€—https://huggingface.co/BRlkl/GRPO-4_70 
Base Model(s)  BRlkl/orchestrator-qwen3-4b-full   BRlkl/orchestrator-qwen3-4b-full
Model Size4b
Required VRAM8.1 GB
Updated2026-09-23
MaintainerBRlkl
Model Typeqwen3
Model Files  5.0 GB: 1-of-2   3.1 GB: 2-of-2
Supported Languagesen
Model ArchitectureQwen3ForCausalLM
Licenseapache-2.0
Context Length262144
Model Max Length262144
Transformers Version4.57.6
Tokenizer ClassQwen2Tokenizer
Padding Token<|im_end|>
Vocabulary Size151936
Errorsreplace

Best Alternatives to GRPO 4 70

Best Alternatives
Context / RAM
Downloads
Likes
Qwen3 4B Instruct 2507256K / 8.1 GB3109972911
FastContext 1.0 4B SFT256K / 8.1 GB5735357
...2.5 Flash Lite Preview Distill256K / 8.1 GB311
Lightning 4B256K / 8.1 GB136
Qwen3 4B Thinking 2507256K / 8.1 GB291023606
Fable Traces256K / 8.1 GB265210
Voho Saudi Chat 4B256K / 8 GB4681
FastContext 1.0 4B RL256K / 8.1 GB455961
Qwen3 4B Instruct 2507 FP8256K / 5.2 GB110074478
Checkpoint 8000256K / 8.1 GB14530
Note: green Score (e.g. "73.2") means that the model is better than BRlkl/GRPO-4_70.