LLM EXPLORER 61,405 MODELS INDEXED

NemotronH 0.3B A0.3B by inference-optimization

By inference-optimization · 737 downloads

NemotronH 0.3B A0.3B is an open-source language model by inference-optimization. Features: 550b LLM, VRAM: 0.7GB, Context: 256K, License: mit, LLM Explorer Score: 0.33.

Base model:finetune:nvidia/nvi... Base model:nvidia/nvidia-nemot...   Endpoints compatible   Nemotron h   Region:us   Safetensors

NemotronH 0.3B A0.3B Parameters and Internals

LLM NameNemotronH 0.3B A0.3B
Repository πŸ€—https://huggingface.co/inference-optimization/NemotronH-0.3B-A0.3B 
Base Model(s)  ...on 3 Ultra 550B A55B Base BF16   nvidia/NVIDIA-Nemotron-3-Ultra-550B-A55B-Base-BF16
Model Size550b
Required VRAM0.7 GB
Updated2026-09-16
Maintainerinference-optimization
Model Typenemotron_h
Model Files  0.7 GB
Model ArchitectureNemotronHForCausalLM
Licensemit
Context Length262144
Model Max Length262144
Transformers Version5.17.0
Tokenizer ClassTokenizersBackend
Padding Token<|im_end|>
Vocabulary Size131072

Best Alternatives to NemotronH 0.3B A0.3B

Best Alternatives
Context / RAM
Downloads
Likes
...motron 3 Ultra 550B A55B NVFP4256K / 63.1 GB240041277
...emotron 3 Ultra 550B A55B BF16256K / 204.9 GB493863308
...motron 3 Ultra 550B A55B NVFP4256K / 63.1 GB1035
Nemotron 3 Labs Ultra Math SFT256K / 85.9 GB5104
Nemotron 3 Labs Ultra Math RL256K / 204.9 GB3743
...motron 3 Ultra 550B A55B GenRM256K / 219.9 GB266611
...on 3 Ultra 550B A55B Base BF16256K / 209.9 GB100529
Nemotron 3 Ultra 550B A55B256K / 144.3 GB4732
...emotron 3 Ultra 550B A55B Base256K / 209.9 GB232
...DIA Nemotron 3 Ultra 550B A55B256K / 204.9 GB191
Note: green Score (e.g. "73.2") means that the model is better than inference-optimization/NemotronH-0.3B-A0.3B.