LLM EXPLORER 59,420 MODELS INDEXED

Mixtral 8x7B Instruct V0.1 OmniQuantv1 W4a16g128 by ChenMnZ

By ChenMnZ · 12 downloads

Mixtral 8x7B Instruct V0.1 OmniQuantv1 W4a16g128 is an open-source language model by ChenMnZ. Features: 6.5b LLM, VRAM: 24.7GB, Context: 32K, MoE, Instruction-Based, LLM Explorer Score: 0.1.

  Arxiv:2308.13137   Conversational   Endpoints compatible   Instruct   Mixtral   Moe   Region:us   Safetensors   Sharded   Tensorflow

Mixtral 8x7B Instruct V0.1 OmniQuantv1 W4a16g128 Parameters and Internals

Additional Notes 
For detailed usage, refer to the corresponding Jupyter notebook available in the repository.
LLM NameMixtral 8x7B Instruct V0.1 OmniQuantv1 W4a16g128
Repository πŸ€—https://huggingface.co/ChenMnZ/Mixtral-8x7B-Instruct-v0.1-OmniQuantv1-w4a16g128 
Model Size6.5b
Required VRAM24.7 GB
Updated2026-08-25
MaintainerChenMnZ
Model Typemixtral
Instruction-BasedYes
Model Files  5.0 GB: 1-of-5   5.0 GB: 2-of-5   5.0 GB: 3-of-5   5.0 GB: 4-of-5   4.7 GB: 5-of-5
Model ArchitectureMixtralForCausalLM
Context Length32768
Model Max Length32768
Transformers Version4.36.0
Tokenizer ClassLlamaTokenizer
Vocabulary Size32000
Torch Data Typefloat16

Best Alternatives to Mixtral 8x7B Instruct V0.1 OmniQuantv1 W4a16g128

Best Alternatives
Context / RAM
Downloads
Likes
...nstruct V0.1 AQLM 2Bit 1x16 Hf32K / 13.1 GB1519
Note: green Score (e.g. "73.2") means that the model is better than ChenMnZ/Mixtral-8x7B-Instruct-v0.1-OmniQuantv1-w4a16g128.