Instella MoE 16B A3B DPO is an open-source language model by amd. Features: 15.9b LLM, VRAM: 31.9GB, Context: 32K, License: other, MoE, LLM Explorer Score: 0.31.
| LLM Name | Instella MoE 16B A3B DPO |
| Repository 🤗 | https://huggingface.co/amd/Instella-MoE-16B-A3B-DPO |
| Model Size | 15.9b |
| Required VRAM | 31.9 GB |
| Updated | 2026-07-25 |
| Maintainer | amd |
| Model Type | deepseek_v3 |
| Model Files | |
| Supported Languages | en |
| Model Architecture | InstellaMoEForCausalLM |
| License | other |
| Context Length | 32768 |
| Model Max Length | 32768 |
| Transformers Version | 4.57.1 |
| Tokenizer Class | LlamaTokenizerFast |
| Beginning of Sentence Token | <|begin▁of▁sentence|> |
| End of Sentence Token | <|end▁of▁sentence|> |
| Vocabulary Size | 128896 |
| Torch Data Type | bfloat16 |
Best Alternatives |
Context / RAM |
Downloads |
Likes |
|---|---|---|---|
| Instella MoE 16B A3B Think | 32K / 31.9 GB | 59 | 25 |
| Instella MoE 16B A3B SFT | 32K / 31.9 GB | 51 | 2 |
| Instella MoE 16B A3B Pretrain | 4K / 31.7 GB | 70 | 2 |
| Instella MoE 16B A3B Midtrain | 4K / 63.5 GB | 70 | 2 |
🆘 Have you tried this model? Rate its performance. This feedback would greatly assist ML community in identifying the most suitable model for their needs. Your contribution really does make a difference! 🌟