LLM EXPLORER 59,962 MODELS INDEXED

Qwen3.8 Flash Next MLX Serve Mixed 4 8bit by ddalcu

By ddalcu · 2422 downloads

Qwen3.8 Flash Next MLX Serve Mixed 4 8bit is an open-source language model by ddalcu. Features: 21.1b LLM, VRAM: 0.1GB, License: other, Quantized, LLM Explorer Score: 0.38.

  4-bit   8bit Base model:quantized:qwen/qwen... Base model:qwen/qwen3.8-flash-...   Conversational   Mlx   Mlx-serve   Moe   Ngram-embedding   Quantized   Qwen4 exp   Region:us   Safetensors   Sparse-attention

Qwen3.8 Flash Next MLX Serve Mixed 4 8bit Parameters and Internals

LLM NameQwen3.8 Flash Next MLX Serve Mixed 4 8bit
Repository πŸ€—https://huggingface.co/ddalcu/Qwen3.8-Flash-Next-MLX-Serve-mixed-4-8bit 
Base Model(s)  Qwen/Qwen3.8-Flash-Next   Qwen/Qwen3.8-Flash-Next
Model Size21.1b
Required VRAM0.1 GB
Updated2026-08-30
Maintainerddalcu
Model Typeqwen4_exp
Model Files  0.1 GB   0.9 GB   0.6 GB   0.9 GB   0.5 GB   0.1 GB   0.9 GB   0.5 GB   0.9 GB   0.6 GB   0.9 GB   0.6 GB   0.9 GB   0.6 GB   0.9 GB   0.5 GB   0.9 GB   0.6 GB   0.9 GB   0.6 GB   0.9 GB   0.6 GB   0.9 GB   0.5 GB   0.9 GB   0.6 GB   0.9 GB   0.5 GB   1.0 GB   0.9 GB   1.0 GB   0.9 GB   0.6 GB   0.9 GB   0.5 GB   0.9 GB   0.6 GB   0.9 GB   0.6 GB   0.9 GB   0.6 GB   0.9 GB   0.5 GB
Quantization Type8bit
Model ArchitectureQwen4ExpForConditionalGeneration
Licenseother
Transformers Version5.8.0.dev0

Best Alternatives to Qwen3.8 Flash Next MLX Serve Mixed 4 8bit

Best Alternatives
Context / RAM
Downloads
Likes
...Next Uncensored MLX Serve 4bit0K / 0.1 GB21666