LLM EXPLORER 59,962 MODELS INDEXED

Qwen3.8 Flash Next MTPLX Optimized Speed by Youssofal

By Youssofal · 2374 downloads

Qwen3.8 Flash Next MTPLX Optimized Speed is an open-source language model by Youssofal. Features: 24.8b LLM, VRAM: 80.3GB, License: other, LLM Explorer Score: 0.4.

  4-bit   Apple-silicon Base model:quantized:qwen/qwen... Base model:qwen/qwen3.8-flash-...   Chat   Conversational   Local-ai   Macos   Mlx   Moe   Mtp   Mtplx   Multi-token-prediction   Qwen   Qwen3.8-flash-next   Qwen4 exp   Region:us   Safetensors   Sharded   Speculative-decoding   Tensorflow

Qwen3.8 Flash Next MTPLX Optimized Speed Parameters and Internals

LLM NameQwen3.8 Flash Next MTPLX Optimized Speed
Repository πŸ€—https://huggingface.co/Youssofal/Qwen3.8-Flash-Next-MTPLX-Optimized-Speed 
Base Model(s)  Qwen/Qwen3.8-Flash-Next   Qwen/Qwen3.8-Flash-Next
Model Size24.8b
Required VRAM80.3 GB
Updated2026-08-30
MaintainerYoussofal
Model Typeqwen4_exp
Model Files  4.6 GB: 1-of-19   4.3 GB: 2-of-19   4.6 GB: 3-of-19   4.3 GB: 4-of-19   4.6 GB: 5-of-19   4.3 GB: 6-of-19   4.6 GB: 7-of-19   4.3 GB: 8-of-19   4.6 GB: 9-of-19   4.3 GB: 10-of-19   4.6 GB: 11-of-19   4.3 GB: 12-of-19   4.6 GB: 13-of-19   4.3 GB: 14-of-19   4.6 GB: 15-of-19   4.3 GB: 16-of-19   4.8 GB: 17-of-19   4.3 GB: 18-of-19   0.0 GB: 19-of-19   0.9 GB   1.7 GB   32.0 GB
Model ArchitectureQwen4ExpForConditionalGeneration
Licenseother
Model Max Length262144
Transformers Version5.8.0.dev0
Tokenizer ClassQwen2Tokenizer
Padding Token<|endoftext|>
Errorsreplace

Best Alternatives to Qwen3.8 Flash Next MTPLX Optimized Speed

Best Alternatives
Context / RAM
Downloads
Likes
...8 Flash Next REAP 384 MLX 4bit0K / 38 GB12875