LLM EXPLORER 58,425 MODELS INDEXED

NVIDIA Nemotron 3.5 Lightning 30B A3B 4bit by mlx-community

By mlx-community · 2504 downloads

NVIDIA Nemotron 3.5 Lightning 30B A3B 4bit is an open-source language model by mlx-community. Features: 30b LLM, VRAM: 17.8GB, Context: 256K, License: other, Quantized, LLM Explorer Score: 0.35.

  4-bit   4bit Base model:nvidia/nvidia-nemot... Base model:quantized:nvidia/nv...   Conversational Dataset:nvidia/nemotron-post-t... Dataset:nvidia/nemotron-pre-tr...   De   En   Es   Fr   It   Ja   Mlx   Nemotron-3.5   Nemotron h   Nvidia   Pytorch   Quantized   Region:us   Safetensors   Sharded   Tensorflow

NVIDIA Nemotron 3.5 Lightning 30B A3B 4bit Parameters and Internals

LLM NameNVIDIA Nemotron 3.5 Lightning 30B A3B 4bit
Repository πŸ€—https://huggingface.co/mlx-community/NVIDIA-Nemotron-3.5-Lightning-30B-A3B-4bit 
Base Model(s)  nvidia/NVIDIA-Nemotron-3.5-Lightning-30B-A3B-BF16   nvidia/NVIDIA-Nemotron-3.5-Lightning-30B-A3B-BF16
Model Size30b
Required VRAM17.8 GB
Updated2026-08-14
Maintainermlx-community
Model Typenemotron_h
Model Files  5.1 GB: 1-of-4   5.3 GB: 2-of-4   5.3 GB: 3-of-4   2.1 GB: 4-of-4
Supported Languagesen es fr de it ja
Quantization Type4bit
Model ArchitectureNemotronHForCausalLM
Licenseother
Context Length262144
Model Max Length262144
Transformers Version4.57.6
Tokenizer ClassTokenizersBackend
Padding Token<|im_end|>
Vocabulary Size131072

Best Alternatives to NVIDIA Nemotron 3.5 Lightning 30B A3B 4bit

Best Alternatives
Context / RAM
Downloads
Likes
...motron 3 Nano 30B A3B MLX 4Bit256K / 17.8 GB2051
...motron 3 Nano 30B A3B MLX 6Bit256K / 25.8 GB1482
...3.5 Lightning 30B A3B MLX 6bit256K / 25.8 GB1821
...tron 3 Nano 30B A3B OptiQ 4bit256K / 22.1 GB6442
...emotron Cascade 2 30B A3B 4bit256K / 17.8 GB84319
...emotron Cascade 2 30B A3B 8bit256K / 33.5 GB2798
...emotron Cascade 2 30B A3B 6bit256K / 25.8 GB2716
...ron Cascade 2 30B A3B Mlx 6bit256K / 27.6 GB983
...A Nemotron 3 Nano 30B A3B 4bit256K / 17.8 GB7284
...emotron Cascade 2 30B A3B 5bit256K / 21.6 GB201