Consensuslab Peer Deference Agreement Qwen2.5 3B Grpo Lora is an open-source language model by tanny2109. Features: 3b LLM, Instruction-Based, LLM Explorer Score: 0.27.
| LLM Name | Consensuslab Peer Deference Agreement Qwen2.5 3B Grpo Lora |
| Repository π€ | https://huggingface.co/tanny2109/consensuslab-peer-deference-agreement-qwen2.5-3b-grpo-lora |
| Base Model(s) | |
| Model Size | 3b |
| Required VRAM | 0 GB |
| Updated | 2026-10-11 |
| Maintainer | tanny2109 |
| Instruction-Based | Yes |
| Model Files | |
| Model Architecture | Adapter |
| Model Max Length | 131072 |
| Is Biased | none |
| Tokenizer Class | Qwen2Tokenizer |
| Padding Token | <|im_end|> |
| PEFT Type | LORA |
| LoRA Model | Yes |
| PEFT Target Modules | k_proj|v_proj|q_proj|o_proj |
| LoRA Alpha | 16 |
| LoRA Dropout | 0 |
| R Param | 8 |
| Errors | replace |
Best Alternatives |
Context / RAM |
Downloads |
Likes |
|---|---|---|---|
| AIBio LoRA | 0K / 0.1 GB | 28 | 1 |
| StandardOne 3B LoRA | 0K / 0.1 GB | 42 | 1 |
| ... Checking Qwen2.5 3B Grpo Lora | 0K / 0 GB | 9 | 0 |
| Qwen2.5 VL 3B FakeBench LoRA | 0K / 0.1 GB | 19 | 1 |
| Talvion Support Sft Kb V2 Lora | 0K / 0.1 GB | 10 | 0 |
| ...2.5 3B Instruct Sheldon SFT V2 | 0K / 0.2 GB | 9 | 1 |
| Running Coach Qwen3b Lora | 0K / 0.1 GB | 27 | 1 |
| ...wen2.5 Vl 3B Food Extract Lora | 0K / 0.1 GB | 15 | 1 |
| Qwen Darija Domain Adapted | 0K / 0.1 GB | 13 | 1 |
| Qwen Mom Generator | 0K / 0.1 GB | 11 | 1 |