LLM EXPLORER 61,755 MODELS INDEXED

Mixtral Instruct Seqlen 4096 Bs 4 Optimum 0 0 23 by aws-neuron

By aws-neuron · 9 downloads

Mixtral Instruct Seqlen 4096 Bs 4 Optimum 0 0 23 is an open-source language model by aws-neuron. Features: LLM, Context: 32K, MoE, Instruction-Based, LLM Explorer Score: 0.13.

  Conversational   Endpoints compatible   Instruct   Mixtral   Moe   Region:us

Mixtral Instruct Seqlen 4096 Bs 4 Optimum 0 0 23 Parameters and Internals

LLM NameMixtral Instruct Seqlen 4096 Bs 4 Optimum 0 0 23
Repository πŸ€—https://huggingface.co/aws-neuron/mixtral-instruct-seqlen-4096-bs-4-optimum-0-0-23 
Updated2026-08-07
Maintaineraws-neuron
Model Typemixtral
Instruction-BasedYes
Model ArchitectureMixtralForCausalLM
Context Length32768
Model Max Length32768
Transformers Version4.41.1
Tokenizer ClassLlamaTokenizer
Vocabulary Size32000
Torch Data Typebfloat16

Best Alternatives to Mixtral Instruct Seqlen 4096 Bs 4 Optimum 0 0 23

Best Alternatives
Context / RAM
Downloads
Likes
...ixtral 8x22B Instruct V0.1 FP864K / 140.9 GB980
Dolphin 2.6 Mixtral 8x7b32K / 93.6 GB11285211
Dolphin 2.6 Mixtral 8x7b32K / 93.6 GB7806217
...eqlen 4096 Bs 4 Optimum 0 0 2332K /  GB80
Dolphin 2.7 Mixtral 8x7b32K / 93.6 GB2178170
Empower Functions Medium32K / 93.6 GB171
Mixtral 8x7B Instruct V0.132K /  GB50
...ral 8x7b Instruct V0.1 Int4 Ov32K / 0 GB164
Dolphin 2.7 Mixtral 8x7b32K / 93.6 GB228169
...ct V0.1 Agent Function Calling32K / 44.3 GB42
Note: green Score (e.g. "73.2") means that the model is better than aws-neuron/mixtral-instruct-seqlen-4096-bs-4-optimum-0-0-23.