LLM EXPLORER 59,420 MODELS INDEXED

LLaMA3 Iterative DPO Final GGUF by sirovub

By sirovub · 13 downloads

LLaMA3 Iterative DPO Final GGUF is an open-source language model by sirovub. Features: LLM, VRAM: 16.1GB, Context: 8K, License: llama3, Quantized, LLM Explorer Score: 0.13.

  Arxiv:2312.11456   Arxiv:2405.07863   Conversational   Endpoints compatible   Gguf   Llama   Q8   Quantized   Region:us   Sharded   Tensorflow

LLaMA3 Iterative DPO Final GGUF Parameters and Internals

Model Type 
text generation, instruction-following
Additional Notes 
RLHFlow\LLaMA3-iterative-DPO-final is an unofficial checkpoint developed for research purposes. While safety and ethical considerations are integral to the alignment process, there remains the possibility that the model could generate offensive or unethical content under adversarial conditions.
Training Details 
Data Sources:
https://huggingface.co/datasets/hendrydong/preference_700K, https://huggingface.co/datasets/RLHFlow/prompt-collection-v0.1
Methodology:
Iterative DPO
LLM NameLLaMA3 Iterative DPO Final GGUF
Repository πŸ€—https://huggingface.co/sirovub/LLaMA3-iterative-DPO-final-GGUF 
Required VRAM16.1 GB
Updated2026-08-04
Maintainersirovub
Model Typellama
Model Files  8.5 GB   16.1 GB   5.0 GB: 1-of-4   5.0 GB: 2-of-4   4.9 GB: 3-of-4   1.2 GB: 4-of-4
GGUF QuantizationYes
Quantization Typegguf|q8
Model ArchitectureLlamaForCausalLM
Licensellama3
Context Length8192
Model Max Length8192
Transformers Version4.40.0.dev0
Tokenizer ClassPreTrainedTokenizerFast
Padding Token<|end_of_text|>
Vocabulary Size128256
Torch Data Typebfloat16

Best Alternatives to LLaMA3 Iterative DPO Final GGUF

Best Alternatives
Context / RAM
Downloads
Likes
KernelLLM GGUF128K / 2.2 GB4542
LLAMA2 GOOD GGUF16K / 4.8 GB150
Codellama Cairo Instruct GGUF16K / 4.1 GB321
Aware Ai 1st8K / 16.1 GB300
MFANNv0.6 GGUF8K / 4.7 GB140
UlizaLlama Q4 K M Gguf4K / 4.2 GB760
Tinyllama Coder Py V154K / 0.7 GB980
Tinyllama Coder Py V164K / 0.7 GB120
Cancer Llama.5 Llm4K / 4.1 GB250
Airavata GGUF4K / 4.2 GB432
Note: green Score (e.g. "73.2") means that the model is better than sirovub/LLaMA3-iterative-DPO-final-GGUF.