LLM EXPLORER 61,755 MODELS INDEXED

Llama 3.1 Nemotron 70B Instruct HF FP8 Dynamic by neuralmagic

By neuralmagic · 436 downloads

Llama 3.1 Nemotron 70B Instruct HF FP8 Dynamic is an open-source language model by neuralmagic. Features: 70b LLM, VRAM: 72.7GB, Context: 128K, License: llama3.1, Instruction-Based, LLM Explorer Score: 0.15.

Base model:nvidia/llama-3.1-ne... Base model:quantized:nvidia/ll...   Compressed-tensors   Conversational   En   Fp8   Instruct   Llama   Region:us   Safetensors   Sharded   Tensorflow   Vllm

Llama 3.1 Nemotron 70B Instruct HF FP8 Dynamic Parameters and Internals

LLM NameLlama 3.1 Nemotron 70B Instruct HF FP8 Dynamic
Repository πŸ€—https://huggingface.co/RedHatAI/Llama-3.1-Nemotron-70B-Instruct-HF-FP8-dynamic 
Base Model(s)  nvidia/Llama-3.1-Nemotron-70B-Instruct-HF   nvidia/Llama-3.1-Nemotron-70B-Instruct-HF
Model Size70b
Required VRAM72.7 GB
Updated2026-08-08
Maintainerneuralmagic
Model Typellama
Instruction-BasedYes
Model Files  4.8 GB: 1-of-15   5.0 GB: 2-of-15   4.9 GB: 3-of-15   4.9 GB: 4-of-15   4.9 GB: 5-of-15   5.0 GB: 6-of-15   4.9 GB: 7-of-15   4.9 GB: 8-of-15   4.9 GB: 9-of-15   5.0 GB: 10-of-15   4.9 GB: 11-of-15   4.9 GB: 12-of-15   4.9 GB: 13-of-15   5.0 GB: 14-of-15   3.8 GB: 15-of-15
Supported Languagesen
Model ArchitectureLlamaForCausalLM
Licensellama3.1
Context Length131072
Model Max Length131072
Transformers Version4.45.0
Tokenizer ClassPreTrainedTokenizerFast
Vocabulary Size128256
Torch Data Typebfloat16

Best Alternatives to Llama 3.1 Nemotron 70B Instruct HF FP8 Dynamic

Best Alternatives
Context / RAM
Downloads
Likes
... Chat 1048K Chinese Llama3 70B1024K / 141.9 GB90695
... Chat 1048K Chinese Llama3 70B1024K / 141.9 GB76844
... 3 70B Instruct Gradient 1048K1024K / 141.9 GB22122
Llama3 Function Calling 1048K1024K / 141.9 GB51
...a 3 70B Instruct Gradient 524K512K / 141.9 GB2423
...a 3 70B Instruct Gradient 262K256K / 141.9 GB2056
...ama 3 70B Arimas Story RP V2.0256K / 141.1 GB103
...ama 3 70B Arimas Story RP V1.6256K / 141.2 GB50
...ama 3 70B Arimas Story RP V1.5256K / 141.2 GB213
Llama 3.1 70B Instruct128K / 141.9 GB813492938
Note: green Score (e.g. "73.2") means that the model is better than RedHatAI/Llama-3.1-Nemotron-70B-Instruct-HF-FP8-dynamic.