Qwen3.8 Flash Next MLX Serve Mixed 4 8bit is an open-source language model by ddalcu. Features: 21.1b LLM, VRAM: 0.1GB, License: other, Quantized, LLM Explorer Score: 0.38.
| LLM Name | Qwen3.8 Flash Next MLX Serve Mixed 4 8bit |
| Repository π€ | https://huggingface.co/ddalcu/Qwen3.8-Flash-Next-MLX-Serve-mixed-4-8bit |
| Base Model(s) | |
| Model Size | 21.1b |
| Required VRAM | 0.1 GB |
| Updated | 2026-08-30 |
| Maintainer | ddalcu |
| Model Type | qwen4_exp |
| Model Files | |
| Quantization Type | 8bit |
| Model Architecture | Qwen4ExpForConditionalGeneration |
| License | other |
| Transformers Version | 5.8.0.dev0 |
Best Alternatives |
Context / RAM |
Downloads |
Likes |
|---|---|---|---|
| ...Next Uncensored MLX Serve 4bit | 0K / 0.1 GB | 2166 | 6 |