LLM EXPLORER 58,297 MODELS INDEXED

Nemotron Cascade 2 30B A3B 5bit by mlx-community

By mlx-community · 20 downloads

Nemotron Cascade 2 30B A3B 5bit is an open-source language model by mlx-community. Features: 30b LLM, VRAM: 21.6GB, Context: 256K, License: other, Quantized, LLM Explorer Score: 0.23.

  5-bit   5bit Base model:nvidia/nemotron-cas... Base model:quantized:nvidia/ne...   Conversational   Custom code   En   General-purpose   Mlx   Nemotron-cascade-2   Nemotron h   Nvidia   Quantized   Reasoning   Region:us   Rl   Safetensors   Sft   Sharded   Tensorflow

Nemotron Cascade 2 30B A3B 5bit Parameters and Internals

LLM NameNemotron Cascade 2 30B A3B 5bit
Repository πŸ€—https://huggingface.co/mlx-community/Nemotron-Cascade-2-30B-A3B-5bit 
Base Model(s)  Nemotron Cascade 2 30B A3B   nvidia/Nemotron-Cascade-2-30B-A3B
Model Size30b
Required VRAM21.6 GB
Updated2026-08-10
Maintainermlx-community
Model Typenemotron_h
Model Files  5.3 GB: 1-of-5   5.1 GB: 2-of-5   5.0 GB: 3-of-5   5.1 GB: 4-of-5   1.1 GB: 5-of-5
Supported Languagesen
Quantization Type5bit
Model ArchitectureNemotronHForCausalLM
Licenseother
Context Length262144
Model Max Length262144
Transformers Version4.55.4
Tokenizer ClassTokenizersBackend
Padding Token<|im_end|>
Vocabulary Size131072
Torch Data Typebfloat16

Best Alternatives to Nemotron Cascade 2 30B A3B 5bit

Best Alternatives
Context / RAM
Downloads
Likes
...motron 3 Nano 30B A3B MLX 4Bit256K / 17.8 GB2051
...motron 3 Nano 30B A3B MLX 6Bit256K / 25.8 GB1482
...3.5 Lightning 30B A3B MLX 6bit256K / 25.8 GB1821
...tron 3 Nano 30B A3B OptiQ 4bit256K / 22.1 GB6442
...emotron Cascade 2 30B A3B 4bit256K / 17.8 GB84319
...emotron Cascade 2 30B A3B 8bit256K / 33.5 GB2798
...emotron Cascade 2 30B A3B 6bit256K / 25.8 GB2716
...ron Cascade 2 30B A3B Mlx 6bit256K / 27.6 GB983
...A Nemotron 3 Nano 30B A3B 4bit256K / 17.8 GB7284
...motron 3 Nano 30B A3B MLX 8Bit256K / 33.5 GB2392
Note: green Score (e.g. "73.2") means that the model is better than mlx-community/Nemotron-Cascade-2-30B-A3B-5bit.