LLM EXPLORER 57,918 MODELS INDEXED

GRPO 4 70 by BRlkl

By BRlkl · 7 downloads

GRPO 4 70 is an open-source language model by BRlkl. Features: 4b LLM, VRAM: 8.1GB, Context: 256K, License: apache-2.0, LLM Explorer Score: 0.37, ELO: 1409.

Base model:brlkl/orchestrator-... Base model:finetune:brlkl/orch...   Conversational   En   Endpoints compatible   Grpo   Qwen3   Region:us   Safetensors   Sharded   Tensorflow   Trl   Unsloth
Model Card on HF πŸ€—: https://huggingface.co/BRlkl/GRPO-4_70 

GRPO 4 70 Benchmarks

nn.n% — How the model compares to the reference models: Anthropic Sonnet 3.5 ("so35"), GPT-4o ("gpt4o") or GPT-4 ("gpt4").

GRPO 4 70 Parameters and Internals

LLM NameGRPO 4 70
Repository πŸ€—https://huggingface.co/BRlkl/GRPO-4_70 
Base Model(s)  BRlkl/orchestrator-qwen3-4b-full   BRlkl/orchestrator-qwen3-4b-full
Model Size4b
Required VRAM8.1 GB
Updated2026-08-10
MaintainerBRlkl
Model Typeqwen3
Model Files  5.0 GB: 1-of-2   3.1 GB: 2-of-2
Supported Languagesen
Model ArchitectureQwen3ForCausalLM
Licenseapache-2.0
Context Length262144
Model Max Length262144
Transformers Version4.57.6
Tokenizer ClassQwen2Tokenizer
Padding Token<|im_end|>
Vocabulary Size151936
Errorsreplace

Best Alternatives to GRPO 4 70

Best Alternatives
Context / RAM
Downloads
Likes
FastContext 1.0 4B SFT256K / 8.1 GB5735357
Qwen3 4B Instruct 2507256K / 8.1 GB3109972911
Fable Traces256K / 8.1 GB578210
Lightning 4B256K / 8.1 GB136
FastContext 1.0 4B RL256K / 8.1 GB455961
Qwen3 4B Thinking 2507256K / 8.1 GB291023606
Qwen3 4B Instruct 2507 FP8256K / 5.2 GB110074478
Typhoon2.5 Qwen3 4B256K / 8 GB4898776
Neuron 4B Instruct256K / 8.1 GB3131
Tac 1256K / 8.1 GB8080
Note: green Score (e.g. "73.2") means that the model is better than BRlkl/GRPO-4_70.