LLM EXPLORER 59,420 MODELS INDEXED

DPO Llama3 8B Grammar Rules by hannahbillo

By hannahbillo · 5 downloads

DPO Llama3 8B Grammar Rules is an open-source language model by hannahbillo. Features: 8b LLM, VRAM: 0.1GB, License: llama3.1, LLM Explorer Score: 0.14.

  Adapter Base model:adapter:meta-llama/... Base model:meta-llama/llama-3....   Dpo   Finetuned   Generated from trainer   Lora   Peft   Region:us   Safetensors   Tensorboard   Trl

DPO Llama3 8B Grammar Rules Parameters and Internals

LLM NameDPO Llama3 8B Grammar Rules
Repository πŸ€—https://huggingface.co/hannahbillo/dpo-llama3-8b-grammar-rules 
Base Model(s)  meta-llama/Meta-Llama-3.1-8B   meta-llama/Meta-Llama-3.1-8B
Model Size8b
Required VRAM0.1 GB
Updated2026-08-04
Maintainerhannahbillo
Model Files  0.1 GB   0.0 GB
Model ArchitectureAdapter
Licensellama3.1
Model Max Length131072
Is Biasednone
Tokenizer ClassPreTrainedTokenizerFast
Padding Token<|end_of_text|>
PEFT TypeLORA
LoRA ModelYes
PEFT Target Modulesq_proj|up_proj|v_proj|o_proj|k_proj|gate_proj|down_proj
LoRA Alpha32
LoRA Dropout0.05
R Param6

Best Alternatives to DPO Llama3 8B Grammar Rules

Best Alternatives
Context / RAM
Downloads
Likes
... 3 8B Instruct Bvr Finetune V38K / 16.1 GB60
...Max H3 Prompt Rewriter LoRA 8B0K / 2.8 GB13726
Fovis0K / 0.1 GB391
Welles0K / 0 GB91
Drv Corrected Opd Qwen3 Vl 8B0K / 0.3 GB61
...eta Llama 3 8B Instruct Amazon0K / 0 GB60
Victor Triage Lora Llama3.1 8B0K / 0.1 GB70
Flippa V60K / 0 GB81
Llama 3 Korean 8B R V 0.10K / 0 GB50
...B Instruct DPO 0R100L PoliTune0K / 16.1 GB50
Note: green Score (e.g. "73.2") means that the model is better than hannahbillo/dpo-llama3-8b-grammar-rules.