LLM EXPLORER 57,918 MODELS INDEXED

Toxicity Reward Model P8 V Head Prompt Output Max Margin Seed 42 Llama 3.2 1B Final by ajagota71

By ajagota71 · 9 downloads

Toxicity Reward Model P8 V Head Prompt Output Max Margin Seed 42 Llama 3.2 1B Final is an open-source language model by ajagota71. Features: 1b LLM, VRAM: 4.9GB, Context: 128K, LLM Explorer Score: 0.19.

  Irl   Llama   Region:us   Reward-model   Safetensors   Toxicity

Toxicity Reward Model P8 V Head Prompt Output Max Margin Seed 42 Llama 3.2 1B Final Parameters and Internals

LLM NameToxicity Reward Model P8 V Head Prompt Output Max Margin Seed 42 Llama 3.2 1B Final
Repository πŸ€—https://huggingface.co/ajagota71/toxicity-reward-model-p8-v-head-prompt-output-max-margin-seed-42-llama-3.2-1b-final 
Model Size1b
Required VRAM4.9 GB
Updated2026-07-01
Maintainerajagota71
Model Typellama
Model Files  4.9 GB   0.0 GB
Supported Languagesen
Model ArchitectureLlamaForCausalLM
Context Length131072
Model Max Length131072
Transformers Version4.54.0
Tokenizer ClassPreTrainedTokenizerFast
Padding Token<|end_of_text|>
Vocabulary Size128256
Torch Data Typefloat32

Best Alternatives to Toxicity Reward Model P8 V Head Prompt Output Max Margin Seed 42 Llama 3.2 1B Final

Best Alternatives
Context / RAM
Downloads
Likes
ISA 02 Nano Llama 3.2 1B1024K / 2.5 GB473
LWM Text 1M1024K / 13.5 GB114929
LWM Text Chat 1M1024K / 13.5 GB57173
JOSIE 1M Base1024K / 13.5 GB121
JOSIE 1M Base1024K / 13.5 GB61
MiniCPM5 1B128K / 2.2 GB354385868
...1B Claude Opus Fable5 Thinking128K / 2.2 GB6165166
Llama 3.2 1B Instruct128K / 2.5 GB212072100
Llama 3.2 1B Instruct128K / 2.5 GB101146341549
LlaMa3.2 1B Instruct128K / 4.9 GB140
Note: green Score (e.g. "73.2") means that the model is better than ajagota71/toxicity-reward-model-p8-v-head-prompt-output-max-margin-seed-42-llama-3.2-1b-final.