LLM EXPLORER 58,133 MODELS INDEXED

NVIDIA Nemotron 3.5 Lightning 30B A3B NVFP4 DFlash by nvidia

By nvidia · 112 downloads

NVIDIA Nemotron 3.5 Lightning 30B A3B NVFP4 DFlash is an open-source language model by nvidia. Features: 30b LLM, VRAM: 1.2GB, Context: 1024K, License: other, LLM Explorer Score: 0.31.

  Arxiv:2602.06036 Base model:nvidia/nvidia-nemot... Base model:quantized:nvidia/nv...   Dflash   Latent-moe   Model optimizer   Modelopt   Mtp   Nemotron-3.5-lightning   Nvidia   Qwen3   Region:us   Safetensors

NVIDIA Nemotron 3.5 Lightning 30B A3B NVFP4 DFlash Parameters and Internals

LLM NameNVIDIA Nemotron 3.5 Lightning 30B A3B NVFP4 DFlash
Repository πŸ€—https://huggingface.co/nvidia/NVIDIA-Nemotron-3.5-Lightning-30B-A3B-NVFP4-DFlash 
Base Model(s)  nvidia/NVIDIA-Nemotron-3.5-Lightning-30B-A3B-BF16   nvidia/NVIDIA-Nemotron-3.5-Lightning-30B-A3B-NVFP4   nvidia/NVIDIA-Nemotron-3.5-Lightning-30B-A3B-BF16   nvidia/NVIDIA-Nemotron-3.5-Lightning-30B-A3B-NVFP4
Model Size30b
Required VRAM1.2 GB
Updated2026-08-11
Maintainernvidia
Model Typeqwen3
Model Files  1.2 GB
Model ArchitectureDFlashDraftModel
Licenseother
Context Length1048576
Model Max Length1048576
Transformers Version5.5.3
Vocabulary Size131072

Best Alternatives to NVIDIA Nemotron 3.5 Lightning 30B A3B NVFP4 DFlash

Best Alternatives
Context / RAM
Downloads
Likes
Qwen3 Coder 30B A3B DFlash256K / 0.9 GB264534