LLM EXPLORER 63,407 MODELS INDEXED

DPO Llama3 8B Grammar Rules by hannahbillo

By hannahbillo · 5 downloads

DPO Llama3 8B Grammar Rules is an open-source language model by hannahbillo. Features: 8b LLM, VRAM: 0.1GB, License: llama3.1, LLM Explorer Score: 0.13.

  Adapter Base model:adapter:meta-llama/... Base model:meta-llama/llama-3....   Dpo   Finetuned   Generated from trainer   Lora   Peft   Region:us   Safetensors   Tensorboard   Trl

DPO Llama3 8B Grammar Rules Parameters and Internals

LLM NameDPO Llama3 8B Grammar Rules
Repository πŸ€—https://huggingface.co/hannahbillo/dpo-llama3-8b-grammar-rules 
Base Model(s)  meta-llama/Meta-Llama-3.1-8B   meta-llama/Meta-Llama-3.1-8B
Model Size8b
Required VRAM0.1 GB
Updated2026-08-04
Maintainerhannahbillo
Model Files  0.1 GB   0.0 GB
Model ArchitectureAdapter
Licensellama3.1
Model Max Length131072
Is Biasednone
Tokenizer ClassPreTrainedTokenizerFast
Padding Token<|end_of_text|>
PEFT TypeLORA
LoRA ModelYes
PEFT Target Modulesq_proj|up_proj|v_proj|o_proj|k_proj|gate_proj|down_proj
LoRA Alpha32
LoRA Dropout0.05
R Param6

Best Alternatives to DPO Llama3 8B Grammar Rules

Best Alternatives
Context / RAM
Downloads
Likes
... 3 8B Instruct Bvr Finetune V38K / 16.1 GB110
...Max H3 Prompt Rewriter LoRA 8B0K / 2.8 GB13726
Seshat Ohada 8B V0.2.0 Beta80K / 0.3 GB181
ProactiveInquirer Qwen3 8B0K / 0.3 GB332
...27 Llama31 8B Sgd New Rank 1280K / 1.3 GB201
VitaGridProtocol0K /  GB151
Qwen3 8B Text2sql Qlora0K / 0.2 GB411
...ntic Coder Abliterated V3 LoRA0K / 0.1 GB101
Triage Qwen3 8B Lora0K / 0.2 GB131
...alized Harm V1 Qwen3 8B Seed430K / 0.2 GB101
Note: green Score (e.g. "73.2") means that the model is better than hannahbillo/dpo-llama3-8b-grammar-rules.