LLM EXPLORER 59,420 MODELS INDEXED

Policy Iteration 1 by AngelRaychev

By AngelRaychev · 86 downloads

Policy Iteration 1 is an open-source language model by AngelRaychev. Features: 134.5m LLM, VRAM: 0.5GB, Context: 8K, License: proprietary, LLM Explorer Score: 0.18.

  Autotrain compatible Base model:angelraychev/policy... Base model:finetune:angelraych...   Conversational   Endpoints compatible   Generated from trainer   Llama   Region:us   Safetensors   Sft   Trl

Policy Iteration 1 Parameters and Internals

LLM NamePolicy Iteration 1
Repository πŸ€—https://huggingface.co/AngelRaychev/policy_iteration_1 
Model Namepolicy_iteration_1
Base Model(s)  Policy Iteration 1   AngelRaychev/policy_iteration_1
Model Size134.5m
Required VRAM0.5 GB
Updated2025-05-24
MaintainerAngelRaychev
Model Typellama
Model Files  0.5 GB   0.0 GB
Gated ModelYes
Model ArchitectureLlamaForCausalLM
Licenseproprietary
Context Length8192
Model Max Length8192
Transformers Version4.51.2
Tokenizer ClassGPT2Tokenizer
Padding Token<|im_end|>
Vocabulary Size49152
Torch Data Typefloat32

Best Alternatives to Policy Iteration 1

Best Alternatives
Context / RAM
Downloads
Likes
BWork LLM8K / 0.3 GB60
HyzeMini8K / 0.3 GB261
Poker SmolLM8K / 0.5 GB16072
SmolLM2 FT MyDataset8K / 0.5 GB50
SmolLM2 FT MyDataset8K / 0.5 GB70
Distilled Chat Ocra8K / 0.5 GB50
SmolLM FT NYTCw2K / 0.5 GB60
N12K / 0.5 GB231
Value Iteration 12K / 0.5 GB120
SmolLM FT CoEdIT2K / 0.5 GB60
Note: green Score (e.g. "73.2") means that the model is better than AngelRaychev/policy_iteration_1.