LLM EXPLORER 58,873 MODELS INDEXED

NVIDIA Nemotron 3.5 Lightning 30B A3B OptiQ 4bit by mlx-community

By mlx-community · 416 downloads

NVIDIA Nemotron 3.5 Lightning 30B A3B OptiQ 4bit is an open-source language model by mlx-community. Features: 30b LLM, VRAM: 22.1GB, Context: 256K, License: other, Quantized, LLM Explorer Score: 0.31.

  4-bit   4bit   Apple-silicon Base model:nvidia/nvidia-nemot... Base model:quantized:nvidia/nv...   Conversational   Mamba   Mixed-precision   Mlx   Moe   Nemotron   Nemotron-h   Nemotron h   Optiq   Quantized   Region:us   Safetensors   Sharded   Tensorflow

NVIDIA Nemotron 3.5 Lightning 30B A3B OptiQ 4bit Parameters and Internals

LLM NameNVIDIA Nemotron 3.5 Lightning 30B A3B OptiQ 4bit
Repository πŸ€—https://huggingface.co/mlx-community/NVIDIA-Nemotron-3.5-Lightning-30B-A3B-OptiQ-4bit 
Base Model(s)  nvidia/NVIDIA-Nemotron-3.5-Lightning-30B-A3B-BF16   nvidia/NVIDIA-Nemotron-3.5-Lightning-30B-A3B-BF16
Model Size30b
Required VRAM22.1 GB
Updated2026-08-18
Maintainermlx-community
Model Typenemotron_h
Model Files  5.4 GB: 1-of-5   5.1 GB: 2-of-5   5.2 GB: 3-of-5   5.3 GB: 4-of-5   1.1 GB: 5-of-5
Quantization Type4bit
Model ArchitectureNemotronHForCausalLM
Licenseother
Context Length262144
Model Max Length262144
Transformers Version4.57.6
Tokenizer ClassTokenizersBackend
Padding Token<|im_end|>
Vocabulary Size131072

Best Alternatives to NVIDIA Nemotron 3.5 Lightning 30B A3B OptiQ 4bit

Best Alternatives
Context / RAM
Downloads
Likes
...motron 3 Nano 30B A3B MLX 4Bit256K / 17.8 GB2051
...motron 3 Nano 30B A3B MLX 6Bit256K / 25.8 GB1482
...ron 3.5 Lightning 30B A3B 4bit256K / 17.8 GB25044
...3.5 Lightning 30B A3B MLX 6bit256K / 25.8 GB1821
...tron 3 Nano 30B A3B OptiQ 4bit256K / 22.1 GB6442
...emotron Cascade 2 30B A3B 4bit256K / 17.8 GB84319
...emotron Cascade 2 30B A3B 8bit256K / 33.5 GB2798
...emotron Cascade 2 30B A3B 6bit256K / 25.8 GB2716
...ron Cascade 2 30B A3B Mlx 6bit256K / 27.6 GB983
...A Nemotron 3 Nano 30B A3B 4bit256K / 17.8 GB7284
Note: green Score (e.g. "73.2") means that the model is better than mlx-community/NVIDIA-Nemotron-3.5-Lightning-30B-A3B-OptiQ-4bit.