LLM EXPLORER 61,561 MODELS INDEXED

Qwen3.8 Flash Next REAP320 OQ3e DWQ MTP Vision MLX by Litwein

By Litwein · 295 downloads

Qwen3.8 Flash Next REAP320 OQ3e DWQ MTP Vision MLX is an open-source language model by Litwein. Features: 133.7b LLM, VRAM: 70.6GB, License: other, Instruction-Based, LLM Explorer Score: 0.32.

  3-bit Base model:quantized:qwen/qwen... Base model:qwen/qwen3.8-flash-...   Conversational   Dataset:agentsea/wave-ui-25k Dataset:allenai/tulu-3-sft-mix... Dataset:bigcode/self-oss-instr... Dataset:detection-datasets/coc...   Dataset:huggingfacem4/chartqa Dataset:huggingfacem4/document...   Dataset:lmms-lab/textvqa Dataset:nvidia/opencodereasoni... Dataset:open-r1/openr1-math-22... Dataset:open-r1/verifiable-cod... Dataset:open-thoughts/openthou... Dataset:swe-bench/swe-smith-tr...   Dwq   En   Expert-pruning   Image-text-to-text   Instruct   M4q   Mixed-precision   Mlx   Moe   Mtp   Multimodal   Omlx   Oq   Quantization   Qwen3.8   Qwen4 exp   Reap   Region:us   Ru   Safetensors   Sharded   Speculative-decoding   Tensorflow   Vision

Qwen3.8 Flash Next REAP320 OQ3e DWQ MTP Vision MLX Parameters and Internals

LLM NameQwen3.8 Flash Next REAP320 OQ3e DWQ MTP Vision MLX
Repository πŸ€—https://huggingface.co/Litwein/Qwen3.8-Flash-Next-REAP320-oQ3e-DWQ-MTP-Vision-MLX 
Base Model(s)  Qwen/Qwen3.8-Flash-Next   Qwen/Qwen3.8-Flash-Next
Model Size133.7b
Required VRAM70.6 GB
Updated2026-09-18
MaintainerLitwein
Model Typeqwen4_exp
Instruction-BasedYes
Model Files  5.2 GB: 1-of-14   5.2 GB: 2-of-14   5.2 GB: 3-of-14   5.2 GB: 4-of-14   5.2 GB: 5-of-14   5.2 GB: 6-of-14   5.2 GB: 7-of-14   5.2 GB: 8-of-14   5.3 GB: 9-of-14   5.3 GB: 10-of-14   5.3 GB: 11-of-14   5.3 GB: 12-of-14   5.3 GB: 13-of-14   2.5 GB: 14-of-14
Supported Languagesen ru
Model ArchitectureQwen4ExpForConditionalGeneration
Licenseother
Model Max Length262144
Transformers Version5.8.0.dev0
Tokenizer ClassQwen2Tokenizer
Padding Token<|endoftext|>
Errorsreplace