LLM EXPLORER 58,425 MODELS INDEXED

NVIDIA Nemotron 3.5 Lightning 30B A3B FP8 by RedHatAI

By RedHatAI · 14 downloads

NVIDIA Nemotron 3.5 Lightning 30B A3B FP8 is an open-source language model by RedHatAI. Features: 30b LLM, VRAM: 32.3GB, Context: 256K, LLM Explorer Score: 0.28.

Base model:nvidia/nvidia-nemot... Base model:quantized:nvidia/nv...   Compressed-tensors   Conversational   Endpoints compatible   Fp8   Llm-compressor   Nemotron-3.5   Nemotron h   Region:us   Safetensors   Sharded   Tensorflow   Vllm

NVIDIA Nemotron 3.5 Lightning 30B A3B FP8 Parameters and Internals

LLM NameNVIDIA Nemotron 3.5 Lightning 30B A3B FP8
Repository πŸ€—https://huggingface.co/RedHatAI/NVIDIA-Nemotron-3.5-Lightning-30B-A3B-FP8 
Base Model(s)  nvidia/NVIDIA-Nemotron-3.5-Lightning-30B-A3B-BF16   nvidia/NVIDIA-Nemotron-3.5-Lightning-30B-A3B-BF16
Model Size30b
Required VRAM32.3 GB
Updated2026-08-14
MaintainerRedHatAI
Model Typenemotron_h
Model Files  20.0 GB: 1-of-2   12.3 GB: 2-of-2
Model ArchitectureNemotronHForCausalLM
Context Length262144
Model Max Length262144
Transformers Version5.15.0
Tokenizer ClassTokenizersBackend
Padding Token<|im_end|>
Vocabulary Size131072

Best Alternatives to NVIDIA Nemotron 3.5 Lightning 30B A3B FP8

Best Alternatives
Context / RAM
Downloads
Likes
...on 3.5 Lightning 30B A3B NVFP41024K / 12.6 GB19250100
...30B A3B NVFP4 Global Pruned 151024K / 11 GB14819
...A Nemotron 3 Nano 30B A3B BF16256K / 63.2 GB1645889752
...ron 3.5 Lightning 30B A3B BF16256K / 65.9 GB1574062
... Nemotron 3 Nano 30B A3B NVFP4256K / 19.3 GB15688
Nemotron Cascade 2 30B A3B256K / 63.2 GB88240518
...IA Nemotron 3 Nano 30B A3B FP8256K / 32.7 GB566152356
... Nemotron 3 Nano 30B A3B NVFP4256K / 19.3 GB765499174
...Nemotron 3.5 Lightning 30B A3B256K / 65.9 GB4744
....5 Lightning 30B A3B Base BF16256K / 65.9 GB33612
Note: green Score (e.g. "73.2") means that the model is better than RedHatAI/NVIDIA-Nemotron-3.5-Lightning-30B-A3B-FP8.