LLM EXPLORER 61,836 MODELS INDEXED

Qwen3.8 Flash Next 119B A5B Niwaki V2.4 3bit Mlx by neopolita

By neopolita · 380 downloads

Qwen3.8 Flash Next 119B A5B Niwaki V2.4 3bit Mlx is an open-source language model by neopolita. Features: 119b LLM, VRAM: 44.8GB, License: other, Quantized, LLM Explorer Score: 0.31.

  4-bit   4bit Base model:mlx-community/qwen3... Base model:quantized:mlx-commu...   Conversational   Mixture-of-experts   Mlx   Moe   Pruning   Quantization   Quantized   Qwen4 exp   Region:us   Safetensors   Sharded   Tensorflow

Qwen3.8 Flash Next 119B A5B Niwaki V2.4 3bit Mlx Parameters and Internals

LLM NameQwen3.8 Flash Next 119B A5B Niwaki V2.4 3bit Mlx
Repository πŸ€—https://huggingface.co/neopolita/Qwen3.8-Flash-Next-119B-A5B-Niwaki-v2.4-3bit-mlx 
Base Model(s)  mlx-community/Qwen3.8-Flash-Next-4bit   mlx-community/Qwen3.8-Flash-Next-4bit
Model Size119b
Required VRAM44.8 GB
Updated2026-09-21
Maintainerneopolita
Model Typeqwen4_exp
Model Files  5.4 GB: 1-of-9   5.3 GB: 2-of-9   5.3 GB: 3-of-9   5.2 GB: 4-of-9   5.3 GB: 5-of-9   5.4 GB: 6-of-9   5.3 GB: 7-of-9   5.3 GB: 8-of-9   2.3 GB: 9-of-9
Quantization Type4bit
Model ArchitectureQwen4ExpForConditionalGeneration
Licenseother
Model Max Length262144
Transformers Version5.8.0.dev0
Tokenizer ClassQwen2Tokenizer
Padding Token<|endoftext|>
Errorsreplace