LLM EXPLORER 60,519 MODELS INDEXED

Nemotron Cascade 2 30B A3B Mxfp4 by mlx-community

By mlx-community · 31 downloads

Nemotron Cascade 2 30B A3B Mxfp4 is an open-source language model by mlx-community. Features: 30b LLM, VRAM: 16.8GB, Context: 256K, License: other, LLM Explorer Score: 0.23.

  4-bit Base model:nvidia/nemotron-cas... Base model:quantized:nvidia/ne...   Conversational   Custom code   En   General-purpose   Mlx   Nemotron-cascade-2   Nemotron h   Nvidia   Reasoning   Region:us   Rl   Safetensors   Sft   Sharded   Tensorflow

Nemotron Cascade 2 30B A3B Mxfp4 Parameters and Internals

LLM NameNemotron Cascade 2 30B A3B Mxfp4
Repository πŸ€—https://huggingface.co/mlx-community/Nemotron-Cascade-2-30B-A3B-mxfp4 
Base Model(s)  Nemotron Cascade 2 30B A3B   nvidia/Nemotron-Cascade-2-30B-A3B
Model Size30b
Required VRAM16.8 GB
Updated2026-08-10
Maintainermlx-community
Model Typenemotron_h
Model Files  5.2 GB: 1-of-4   5.3 GB: 2-of-4   5.4 GB: 3-of-4   0.9 GB: 4-of-4
Supported Languagesen
Model ArchitectureNemotronHForCausalLM
Licenseother
Context Length262144
Model Max Length262144
Transformers Version4.55.4
Tokenizer ClassTokenizersBackend
Padding Token<|im_end|>
Vocabulary Size131072
Torch Data Typebfloat16

Best Alternatives to Nemotron Cascade 2 30B A3B Mxfp4

Best Alternatives
Context / RAM
Downloads
Likes
...on 3.5 Lightning 30B A3B NVFP41024K / 12.6 GB19250100
...30B A3B NVFP4 Global Pruned 151024K / 11 GB14819
...A Nemotron 3 Nano 30B A3B BF16256K / 63.2 GB1645889752
... Nemotron 3 Nano 30B A3B NVFP4256K / 19.3 GB15688
...ron 3.5 Lightning 30B A3B BF16256K / 65.9 GB1574062
Nemotron Cascade 2 30B A3B256K / 63.2 GB88240518
Phonellm Alpha 1256K / 63.2 GB64107
...IA Nemotron 3 Nano 30B A3B FP8256K / 32.7 GB566152356
... Nemotron 3 Nano 30B A3B NVFP4256K / 19.3 GB765499174
...otron 3 Nano 30B A3B Base BF16256K / 63.2 GB90736127
Note: green Score (e.g. "73.2") means that the model is better than mlx-community/Nemotron-Cascade-2-30B-A3B-mxfp4.