LLM EXPLORER 59,071 MODELS INDEXED

Ling 3.0 Flash HybridQuant NVFP4 W4A16 LocalHessian by JasonW2025

By JasonW2025 · 279 downloads

Ling 3.0 Flash HybridQuant NVFP4 W4A16 LocalHessian is an open-source language model by JasonW2025. Features: 65.5b LLM, VRAM: 76.8GB, Context: 256K, License: other, LLM Explorer Score: 0.32.

  8-bit   Bailing hybrid Base model:inclusionai/ling-3.... Base model:quantized:inclusion...   Conversational   Custom code   Local-hessian   Modelopt   Moe   Nvfp4   Quantized   Region:us   Safetensors   Sharded   Tensorflow   Vllm   W4a16

Ling 3.0 Flash HybridQuant NVFP4 W4A16 LocalHessian Parameters and Internals

LLM NameLing 3.0 Flash HybridQuant NVFP4 W4A16 LocalHessian
Repository πŸ€—https://huggingface.co/JasonW2025/Ling-3.0-flash-HybridQuant-NVFP4-W4A16-LocalHessian 
Base Model(s)  inclusionAI/Ling-3.0-flash   inclusionAI/Ling-3.0-flash
Model Size65.5b
Required VRAM76.8 GB
Updated2026-08-20
MaintainerJasonW2025
Model Typebailing_hybrid
Model Files  10.0 GB: 1-of-8   10.0 GB: 2-of-8   10.0 GB: 3-of-8   10.0 GB: 4-of-8   10.0 GB: 5-of-8   10.0 GB: 6-of-8   10.0 GB: 7-of-8   6.8 GB: 8-of-8
Model ArchitectureBailingMoeV3ForCausalLM
Licenseother
Context Length262144
Model Max Length262144
Transformers Version5.14.1
Tokenizer ClassPreTrainedTokenizerFast
Padding Token<|endoftext|>
Vocabulary Size157184