LLM EXPLORER 62,589 MODELS INDEXED

DPO by harshavaishnav

By harshavaishnav · 120 downloads

DPO is an open-source language model by harshavaishnav. Features: 4b LLM, VRAM: 9.3GB, License: apache-2.0, LLM Explorer Score: 0.25.

Base model:adapter:qwen/qwen3....   Base model:qwen/qwen3.5-4b   Conversational Dataset:harshavaishnav/dpo dat...   Dpo   En   Lm-playschool   Lora   Playpen   Qwen   Qwen3 5   Region:us   Safetensors   Sharded   Tensorflow
Model Card on HF πŸ€—: https://huggingface.co/harshavaishnav/DPO 

DPO Parameters and Internals

LLM NameDPO
Repository πŸ€—https://huggingface.co/harshavaishnav/DPO 
Model NameQwen/Qwen3.5-4B
Base Model(s)  Qwen/Qwen3.5-4B   Qwen/Qwen3.5-4B
Model Size4b
Required VRAM9.3 GB
Updated2026-08-08
Maintainerharshavaishnav
Model Typeqwen3_5
Model Files  5.3 GB: 1-of-2   4.0 GB: 2-of-2
Supported Languagesen
Model ArchitectureQwen3_5ForConditionalGeneration
Licenseapache-2.0
Model Max Length262144
Tokenizer ClassQwen2Tokenizer
Padding Token<|endoftext|>
Torch Data Typebfloat16
Errorsreplace

Best Alternatives to DPO

Best Alternatives
Context / RAM
Downloads
Likes
Qwen3.5 4B0K / 9.3 GB6575502801
Intern Decision 4B0K / 0.1 GB33848
NuExtract30K / 0.2 GB441570324
Agents A1 4B0K / 9.1 GB17757872
Tev1 4B Experimental0K / 9.3 GB153916
OneJev 4B0K / 10.4 GB7672
Qwen3.5 4B Base0K / 9.3 GB22781479
APUS OpenJev V1 4B0K / 9.1 GB6654
CogEvol 4B0K / 9.1 GB90727
Fara1.5 4B0K / 9.1 GB493439
Note: green Score (e.g. "73.2") means that the model is better than harshavaishnav/DPO.