LLM EXPLORER 57,252 MODELS INDEXED

Llama 3.1 Nemotron Nano 8B V1 by nvidia

By nvidia · 118588 downloads

Llama 3.1 Nemotron Nano 8B V1 is an open-source language model by nvidia. Features: 8b LLM, VRAM: 16.1GB, Context: 128K, License: other, LLM Explorer Score: 0.26.

  Arxiv:2502.00203   Arxiv:2505.00949   Conversational   Deploy:azure   En   Endpoints compatible   Llama   Llama-3   Nvidia   Pytorch   Region:us   Safetensors   Sharded   Tensorflow

Llama 3.1 Nemotron Nano 8B V1 Parameters and Internals

LLM NameLlama 3.1 Nemotron Nano 8B V1
Repository πŸ€—https://huggingface.co/nvidia/Llama-3.1-Nemotron-Nano-8B-v1 
Model Size8b
Required VRAM16.1 GB
Updated2026-08-08
Maintainernvidia
Model Typellama
Model Files  5.0 GB: 1-of-4   5.0 GB: 2-of-4   4.9 GB: 3-of-4   1.2 GB: 4-of-4
Supported Languagesen
Model ArchitectureLlamaForCausalLM
Licenseother
Context Length131072
Model Max Length131072
Transformers Version4.47.1
Tokenizer ClassPreTrainedTokenizerFast
Vocabulary Size128256
Torch Data Typebfloat16

Quantized Models of the Llama 3.1 Nemotron Nano 8B V1

Model
Likes
Downloads
VRAM
...a 3.1 Nemotron Nano 8B V1 GGUF1126602 GB

Best Alternatives to Llama 3.1 Nemotron Nano 8B V1

Best Alternatives
Context / RAM
Downloads
Likes
...otron 8B UltraLong 4M Instruct4192K / 32.1 GB1135125
UltraLong Thinking4192K / 16.1 GB23
...a 3.1 8B UltraLong 4M Instruct4192K / 32.1 GB17624
...a 3.1 8B UltraLong 2M Instruct2096K / 32.1 GB8759
...otron 8B UltraLong 2M Instruct2096K / 32.1 GB12418
Cthulhu 8B V1.41048K / 16.1 GB1010
...raLong 1M Instruct Abliterated1048K / 32.1 GB49
...a 3.1 8B UltraLong 1M Instruct1048K / 32.1 GB138729
...otron 8B UltraLong 1M Instruct1048K / 32.1 GB70259
...xis Bookwriter Llama3.1 8B Sft1048K / 16.1 GB254