LLM EXPLORER 63,791 MODELS INDEXED

Ling 3.0 Flash Int4 by inclusionAI

By inclusionAI · 1969 downloads

Ling 3.0 Flash Int4 is an open-source language model by inclusionAI. Features: 127.5b LLM, VRAM: 77.6GB, Context: 256K, License: mit, LLM Explorer Score: 0.44, ELO: 1458.

  Bailing hybrid   Compressed-tensors   Conversational   Custom code   Region:us   Safetensors   Sharded   Tensorflow

Ling 3.0 Flash Int4 Benchmarks

nn.n% — How the model compares to the reference models: Anthropic Sonnet 3.5 ("so35"), GPT-4o ("gpt4o") or GPT-4 ("gpt4").

Ling 3.0 Flash Int4 Parameters and Internals

LLM NameLing 3.0 Flash Int4
Repository πŸ€—https://huggingface.co/inclusionAI/Ling-3.0-flash-int4 
Model Size127.5b
Required VRAM77.6 GB
Updated2026-08-08
MaintainerinclusionAI
Model Typebailing_hybrid
Model Files  8.3 GB: 1-of-24   3.0 GB: 2-of-24   3.0 GB: 3-of-24   3.0 GB: 4-of-24   3.0 GB: 5-of-24   3.0 GB: 6-of-24   3.0 GB: 7-of-24   3.0 GB: 8-of-24   3.0 GB: 9-of-24   3.0 GB: 10-of-24   3.0 GB: 11-of-24   3.0 GB: 12-of-24   3.0 GB: 13-of-24   3.0 GB: 14-of-24   3.0 GB: 15-of-24   3.0 GB: 16-of-24   3.0 GB: 17-of-24   3.0 GB: 18-of-24   3.0 GB: 19-of-24   3.0 GB: 20-of-24   3.0 GB: 21-of-24   3.0 GB: 22-of-24   3.0 GB: 23-of-24   3.3 GB: 24-of-24
Model ArchitectureBailingMoeV3ForCausalLM
Licensemit
Context Length262144
Model Max Length262144
Transformers Version4.45.0
Tokenizer ClassPreTrainedTokenizerFast
Padding Token<|endoftext|>
Vocabulary Size157184
Torch Data Typebfloat16

Best Alternatives to Ling 3.0 Flash Int4

Best Alternatives
Context / RAM
Downloads
Likes
Ling 3.0 Flash256K / 254.5 GB4189213
Ling 3.0 Flash Fin256K / 186 GB34259
Ling 3.0 Flash Fp8256K / 129.3 GB139526
Ling 3.0 Flash Base Midtrain256K / 255.9 GB11555
Ling 3.0 Flash Base256K / 255.9 GB11924
Ling 3.0 Flash CIRU IU4256K / 77.6 GB4545
Ling 3.0 Flash Heretic256K / 255 GB2071
Ling 3.0 Flash Fin Int4256K / 58.6 GB471
Ling 3.0 Flash Fin Fp8256K / 94.9 GB381
...0 Flash CIRU Int4 Strix Native256K / 77.6 GB477
Note: green Score (e.g. "73.2") means that the model is better than inclusionAI/Ling-3.0-flash-int4.