LLM EXPLORER 59,071 MODELS INDEXED

Nemotron H 8B Reasoning 128K by nvidia

By nvidia · 1250 downloads

Nemotron H 8B Reasoning 128K is an open-source language model by nvidia. Features: 8b LLM, VRAM: 16.2GB, Context: 128K, License: other, LLM Explorer Score: 0.2.

  Arxiv:2504.03624   Arxiv:2505.00949   Conversational   En   Endpoints compatible   Nvidia   Pytorch   Region:us   Safetensors   Sharded   Tensorflow

Nemotron H 8B Reasoning 128K Parameters and Internals

LLM NameNemotron H 8B Reasoning 128K
Repository πŸ€—https://huggingface.co/nvidia/Nemotron-H-8B-Reasoning-128K 
Model Size8b
Required VRAM16.2 GB
Updated2026-07-11
Maintainernvidia
Model Typenemotron_h
Model Files  5.0 GB: 1-of-4   4.9 GB: 2-of-4   4.9 GB: 3-of-4   1.4 GB: 4-of-4
Supported Languagesen
Model ArchitectureNemotronHForCausalLM
Licenseother
Context Length131072
Model Max Length131072
Transformers Version4.48.0.dev0
Tokenizer ClassPreTrainedTokenizerFast
Vocabulary Size131072
Torch Data Typebfloat16

Best Alternatives to Nemotron H 8B Reasoning 128K

Best Alternatives
Context / RAM
Downloads
Likes
Luciole 8B Instruct 1.1128K / 16.2 GB16794
Luciole 8B Instruct 1.1 NVFP4128K / 6 GB592
...motron H 8B Reasoning 128K FP8128K / 9.2 GB14513
Nemotron H 8B Base 8K8K / 16.2 GB8423258
Note: green Score (e.g. "73.2") means that the model is better than nvidia/Nemotron-H-8B-Reasoning-128K.