LLM EXPLORER 58,297 MODELS INDEXED

Llama 3.1 8B Instruct GRPO Alpaca Combine 100 No KL 42 by KevinG

By KevinG · 5 downloads

Llama 3.1 8B Instruct GRPO Alpaca Combine 100 No KL 42 is an open-source language model by KevinG. Features: 8b LLM, VRAM: 16.1GB, Context: 128K, Instruction-Based, LLM Explorer Score: 0.19.

  Instruct   Llama   Region:us   Safetensors   Sharded   Tensorflow

Llama 3.1 8B Instruct GRPO Alpaca Combine 100 No KL 42 Parameters and Internals

LLM NameLlama 3.1 8B Instruct GRPO Alpaca Combine 100 No KL 42
Repository πŸ€—https://huggingface.co/KevinG/Llama-3.1-8B-Instruct-GRPO-alpaca-combine-100-no-KL-42 
Model Size8b
Required VRAM16.1 GB
Updated2025-09-23
MaintainerKevinG
Model Typellama
Instruction-BasedYes
Model Files  5.0 GB: 1-of-4   5.0 GB: 2-of-4   4.9 GB: 3-of-4   1.2 GB: 4-of-4   0.0 GB
Model ArchitectureLlamaForCausalLM
Context Length131072
Model Max Length131072
Transformers Version4.50.0
Tokenizer ClassPreTrainedTokenizer
Padding Token<|eot_id|>
Vocabulary Size128256
Torch Data Typebfloat16

Best Alternatives to Llama 3.1 8B Instruct GRPO Alpaca Combine 100 No KL 42

Best Alternatives
Context / RAM
Downloads
Likes
...otron 8B UltraLong 4M Instruct4192K / 32.1 GB1135125
UltraLong Thinking4192K / 16.1 GB23
...a 3.1 8B UltraLong 4M Instruct4192K / 32.1 GB17624
...a 3.1 8B UltraLong 2M Instruct2096K / 32.1 GB8759
...otron 8B UltraLong 2M Instruct2096K / 32.1 GB12418
Cthulhu 8B V1.41048K / 16.1 GB1010
...raLong 1M Instruct Abliterated1048K / 32.1 GB49
...a 3.1 8B UltraLong 1M Instruct1048K / 32.1 GB138729
...otron 8B UltraLong 1M Instruct1048K / 32.1 GB70259
Zero Llama 3.1 8B Beta61048K / 16.1 GB71
Note: green Score (e.g. "73.2") means that the model is better than KevinG/Llama-3.1-8B-Instruct-GRPO-alpaca-combine-100-no-KL-42.