LLM EXPLORER 59,071 MODELS INDEXED

Nemotron H 8B Reasoning 128K FP8 by nvidia

By nvidia · 145 downloads

Nemotron H 8B Reasoning 128K FP8 is an open-source language model by nvidia. Features: 8b LLM, VRAM: 9.2GB, Context: 128K, License: other, LLM Explorer Score: 0.19.

  Arxiv:2504.03624   Arxiv:2505.00949   Conversational   En   Endpoints compatible   Nvidia   Pytorch   Region:us   Safetensors   Sharded   Tensorflow

Nemotron H 8B Reasoning 128K FP8 Parameters and Internals

LLM NameNemotron H 8B Reasoning 128K FP8
Repository πŸ€—https://huggingface.co/nvidia/Nemotron-H-8B-Reasoning-128K-FP8 
Model Size8b
Required VRAM9.2 GB
Updated2026-07-31
Maintainernvidia
Model Typenemotron_h
Model Files  5.0 GB: 1-of-2   4.2 GB: 2-of-2
Supported Languagesen
Model ArchitectureNemotronHForCausalLM
Licenseother
Context Length131072
Model Max Length131072
Transformers Version4.51.3
Tokenizer ClassPreTrainedTokenizerFast
Vocabulary Size131072
Torch Data Typebfloat16

Best Alternatives to Nemotron H 8B Reasoning 128K FP8

Best Alternatives
Context / RAM
Downloads
Likes
Luciole 8B Instruct 1.1128K / 16.2 GB16794
Luciole 8B Instruct 1.1 NVFP4128K / 6 GB592
Nemotron H 8B Reasoning 128K128K / 16.2 GB125028
Nemotron H 8B Base 8K8K / 16.2 GB8423258
Note: green Score (e.g. "73.2") means that the model is better than nvidia/Nemotron-H-8B-Reasoning-128K-FP8.