LLM EXPLORER 57,918 MODELS INDEXED

LLaDA2.0 Flash 8bit by mlx-community

By mlx-community · 11 downloads

LLaDA2.0 Flash 8bit is an open-source language model by mlx-community. Features: 102.9b LLM, VRAM: 109.3GB, Context: 32K, License: apache-2.0, Quantized, LLM Explorer Score: 0.21.

  8-bit   8bit Base model:inclusionai/llada2.... Base model:quantized:inclusion...   Conversational   Custom code   Diffusion   Dllm   Llada2 moe   Mlx   Quantized   Region:us   Safetensors   Sharded   Tensorflow   Text generation

LLaDA2.0 Flash 8bit Parameters and Internals

LLM NameLLaDA2.0 Flash 8bit
Repository πŸ€—https://huggingface.co/mlx-community/LLaDA2.0-flash-8bit 
Base Model(s)  inclusionAI/LLaDA2.0-flash   inclusionAI/LLaDA2.0-flash
Model Size102.9b
Required VRAM109.3 GB
Updated2026-08-10
Maintainermlx-community
Model Typellada2_moe
Model Files  4.4 GB: 1-of-24   4.6 GB: 2-of-24   4.6 GB: 3-of-24   4.7 GB: 4-of-24   4.6 GB: 5-of-24   4.6 GB: 6-of-24   4.7 GB: 7-of-24   4.6 GB: 8-of-24   4.6 GB: 9-of-24   4.7 GB: 10-of-24   4.6 GB: 11-of-24   4.6 GB: 12-of-24   4.7 GB: 13-of-24   4.6 GB: 14-of-24   4.6 GB: 15-of-24   4.7 GB: 16-of-24   4.6 GB: 17-of-24   4.6 GB: 18-of-24   4.7 GB: 19-of-24   4.6 GB: 20-of-24   4.6 GB: 21-of-24   4.7 GB: 22-of-24   4.6 GB: 23-of-24   3.0 GB: 24-of-24
Quantization Type8bit
Model ArchitectureLLaDA2MoeModelLM
Licenseapache-2.0
Context Length32768
Model Max Length32768
Transformers Version4.51.0
Tokenizer ClassPreTrainedTokenizerFast
Padding Token<|endoftext|>
Vocabulary Size157184
Torch Data Typebfloat16

Best Alternatives to LLaDA2.0 Flash 8bit

Best Alternatives
Context / RAM
Downloads
Likes
LLaDA2.2 Flash OptiQ 2bit128K / 38.9 GB613
LLaDA2.0 Flash 4bit32K / 57.8 GB183
LLaDA2.0 Flash Preview 4bit16K / 57.8 GB123
LLaDA2.2 Flash128K / 204.5 GB32857
LLaDA2.1 Flash32K / 204.5 GB2982093
LLaDA2.0 Flash32K / 206 GB51869
LLaDA2.0 Flash CAP32K / 206 GB659
LLaDA2.0 Flash Preview16K / 205.9 GB2467
Note: green Score (e.g. "73.2") means that the model is better than mlx-community/LLaDA2.0-flash-8bit.