LLM EXPLORER 61,491 MODELS INDEXED

Yi 34B 200K Rawrr1 LORA DPO Experimental R2 by adamo1139

By adamo1139 · 3 downloads

Yi 34B 200K Rawrr1 LORA DPO Experimental R2 is an open-source language model by adamo1139. Features: 34b LLM, VRAM: 0.1GB, LLM Explorer Score: 0.1.

  4-bit   Adapter   Bitsandbytes   Finetuned   Generated from trainer   Llama   Lora   Peft   Region:us

Yi 34B 200K Rawrr1 LORA DPO Experimental R2 Parameters and Internals

Model Type 
LlamaForCausalLM
Use Cases 
Areas:
Research, Commercial applications
Primary Use Cases:
Instruct training, Generating itineraries
Limitations:
Sequence length limitation of 200 tokens
Additional Notes 
Model is designed to run efficient fine-tuning with reduced VRAM requirements using quantization techniques like 'bitsandbytes'.
Training Details 
Data Sources:
/run/media/.../axolotl/datasets/rawrr_v1/
Methodology:
Trained via DPO technique
Context Length:
200
Hardware Used:
RTX 3090 Ti
Model Architecture:
Llama
LLM NameYi 34B 200K Rawrr1 LORA DPO Experimental R2
Repository πŸ€—https://huggingface.co/adamo1139/Yi-34B-200K-rawrr1-LORA-DPO-experimental-r2 
Model Size34b
Required VRAM0.1 GB
Updated2026-07-27
Maintaineradamo1139
Model Files  0.1 GB
Model ArchitectureAdapter
Model Max Length200000
Is Biasednone
Tokenizer ClassLlamaTokenizer
Padding Token<unk>
PEFT TypeLORA
LoRA ModelYes
PEFT Target Modulesk_proj|down_proj|up_proj|v_proj|q_proj|gate_proj|o_proj
LoRA Alpha8
LoRA Dropout0.05
R Param4

Best Alternatives to Yi 34B 200K Rawrr1 LORA DPO Experimental R2

Best Alternatives
Context / RAM
Downloads
Likes
30B 2e0K / 0 GB50
...34B 200K AEZAKMI RAW 2301 LoRA0K / 0.5 GB01
Airoboros 34B 3.3 Peft0K / 0.5 GB01
Bagel 34B V0.5 Peft0K / 0.5 GB01
34B Beta Onc V10K / 0 GB41
Deacon 34B Qlora Adapter0K / 0.1 GB02
...pmoney 34B 200K Chat Evaluator0K / 2 GB035
...awrr V1 Run1 Experimental LoRA0K / 0.5 GB61
Yi 34B 200K MiniOrca Adapter0K / 0.1 GB01
Yi 34B Alpaca Cot Lora0K / 0.1 GB015
Note: green Score (e.g. "73.2") means that the model is better than adamo1139/Yi-34B-200K-rawrr1-LORA-DPO-experimental-r2.