LLM EXPLORER 58,133 MODELS INDEXED

SmolVLM Instruct DPO by HuggingFaceTB

By HuggingFaceTB · 17 downloads

SmolVLM Instruct DPO is an open-source language model by HuggingFaceTB. Features: 1.7b LLM, VRAM: 0.1GB, License: apache-2.0, Instruction-Based, LLM Explorer Score: 0.16.

  Adapter Base model:adapter:huggingface... Base model:huggingfacetb/smoll...   Conversational Dataset:huggingfaceh4/rlaif-v ...   Dpo   En   Finetuned   Image-text-to-text   Instruct   Lora   Peft   Region:us   Safetensors   Trl

SmolVLM Instruct DPO Parameters and Internals

LLM NameSmolVLM Instruct DPO
Repository πŸ€—https://huggingface.co/HuggingFaceTB/SmolVLM-Instruct-DPO 
Base Model(s)  HuggingFaceTB/SmolLM2-1.7B-Instruct   google/siglip-so400m-patch14-384   HuggingFaceTB/SmolLM2-1.7B-Instruct   google/siglip-so400m-patch14-384
Model Size1.7b
Required VRAM0.1 GB
Updated2026-08-11
MaintainerHuggingFaceTB
Instruction-BasedYes
Model Files  0.1 GB   0.0 GB
Supported Languagesen
Model ArchitectureAdapter
Licenseapache-2.0
Model Max Length16384
Is Biasednone
Tokenizer ClassGPT2Tokenizer
Padding Token<|im_end|>
Vocabulary Size49152
PEFT TypeLORA
LoRA ModelYes
PEFT Target Modulesfc2|up_proj|down_proj|fc1|out_proj|gate_proj|proj|q_proj|k_proj|o_proj|v_proj
LoRA Alpha32
LoRA Dropout0.05
R Param16