LLM EXPLORER 61,561 MODELS INDEXED

Llama 3.1 Nemotron Nano 8B V1 GGUF by unsloth

By unsloth · 2660 downloads

Llama 3.1 Nemotron Nano 8B V1 GGUF is an open-source language model by unsloth. Features: 8b LLM, VRAM: 2.2GB, Context: 128K, License: other, Quantized, LLM Explorer Score: 0.2.

  Arxiv:2502.00203   Arxiv:2505.00949 Base model:nvidia/llama-3.1-ne... Base model:quantized:nvidia/ll...   Conversational   En   Endpoints compatible   Gguf   Llama   Nvidia   Q2   Quantized   Region:us   Unsloth - llama-3 - pytorch

Llama 3.1 Nemotron Nano 8B V1 GGUF Parameters and Internals

LLM NameLlama 3.1 Nemotron Nano 8B V1 GGUF
Repository πŸ€—https://huggingface.co/unsloth/Llama-3.1-Nemotron-Nano-8B-v1-GGUF 
Base Model(s)  nvidia/Llama-3.1-Nemotron-Nano-8B-v1   nvidia/Llama-3.1-Nemotron-Nano-8B-v1
Model Size8b
Required VRAM2.2 GB
Updated2026-05-20
Maintainerunsloth
Model Typellama
Model Files  16.1 GB   4.7 GB   4.5 GB   3.2 GB   3.3 GB   4.0 GB   3.7 GB   4.7 GB   5.1 GB   4.9 GB   4.7 GB   5.7 GB   5.6 GB   6.6 GB   2.3 GB   2.2 GB   3.0 GB   2.5 GB   3.3 GB   3.4 GB   4.2 GB   5.0 GB   5.7 GB   7.3 GB   10.6 GB
Supported Languagesen
GGUF QuantizationYes
Quantization Typegguf|q2|q4_k|q5_k
Model ArchitectureLlamaForCausalLM
Licenseother
Context Length131072
Model Max Length131072
Transformers Version4.52.2
Vocabulary Size128256
Torch Data Typebfloat16

Best Alternatives to Llama 3.1 Nemotron Nano 8B V1 GGUF

Best Alternatives
Context / RAM
Downloads
Likes
10K V61024K / 16.1 GB50
...truct Gradient 1048K IMat GGUF1024K / 2 GB6506
...B Instruct Gradient 1048K GGUF1024K / 3.2 GB4453
Unhinged Llama3 8B 524K512K / 26.5 GB250
Llama 3 8B Instruct 262K GGUF256K / 3.2 GB2082
Dolphin3 Cyber 8B GGUF128K / 3.2 GB59950150
...s Abliterated V2 Base And GGUF128K / 16.1 GB3012
Grok Oss Apollyon 8B128K / 16.1 GB34498
Dolphin3 Cyber 8B GGUF128K / 3.2 GB4561
...3 8B Instruct 128K Jbliterated128K / 4.9 GB4983
Note: green Score (e.g. "73.2") means that the model is better than unsloth/Llama-3.1-Nemotron-Nano-8B-v1-GGUF.