LLM EXPLORER 59,358 MODELS INDEXED

Llama 3 Teal Instruct 2x8B MoE by RDson

By RDson · 8 downloads

Llama 3 Teal Instruct 2x8B MoE is an open-source language model by RDson. Features: 13.7b LLM, VRAM: 27.3GB, Context: 8K, MoE, Instruction-Based, LLM Explorer Score: 0.12.

  2x8b   3   Conversational   Endpoints compatible   Instruct   Llama   Llama 3   Mixtral   Moe   Region:us   Safetensors   Sharded   Tensorflow

Llama 3 Teal Instruct 2x8B MoE Parameters and Internals

Model Type 
Multimodal, Instruction-based, Expert, MoE
Additional Notes 
This is an experimental MoE model combining Meta-Llama-3-8B-Instruct and NVIDIA's Llama3-ChatQA-1.5-8B using Mergekit.
Training Details 
Methodology:
MoE (Mixture of Experts)
LLM NameLlama 3 Teal Instruct 2x8B MoE
Repository πŸ€—https://huggingface.co/RDson/Llama-3-Teal-Instruct-2x8B-MoE 
Model Size13.7b
Required VRAM27.3 GB
Updated2026-05-21
MaintainerRDson
Model Typemixtral
Instruction-BasedYes
Model Files  5.0 GB: 1-of-6   4.9 GB: 2-of-6   5.0 GB: 3-of-6   5.0 GB: 4-of-6   4.9 GB: 5-of-6   2.5 GB: 6-of-6
Model ArchitectureMixtralForCausalLM
Context Length8192
Model Max Length8192
Transformers Version4.41.0.dev0
Tokenizer ClassPreTrainedTokenizerFast
Padding Token<|begin_of_text|>
Vocabulary Size128256
Torch Data Typefloat16

Best Alternatives to Llama 3 Teal Instruct 2x8B MoE

Best Alternatives
Context / RAM
Downloads
Likes
L3.1 Celestial Stone 2x8B128K / 27.3 GB3123
...ma 3 2x8B Instruct MoE 64K Ctx64K / 27.3 GB124
Defne Llama3 2x8B8K / 27.4 GB156
Inixion 2x8B V28K / 27.4 GB62
MoE Llama3 8bx2 Rag8K / 27.3 GB70
Inixion 2x8B8K / 27.5 GB91
Llama 3 Chatty 2x8B8K / 27.3 GB611
FinalFintetuning XVIII 2x8B8K / 27.5 GB114
Llama 3 8Bx2 MoE DPO8K / 27.4 GB41
...lama3 2x8b MoE 41K Experiment18K / 27.3 GB52
Note: green Score (e.g. "73.2") means that the model is better than RDson/Llama-3-Teal-Instruct-2x8B-MoE.