LLM EXPLORER 59,420 MODELS INDEXED

Llama 3.1 Minitron 4B Width Base by nvidia

By nvidia · 5964 downloads

Llama 3.1 Minitron 4B Width Base is an open-source language model by nvidia. Features: 4b LLM, VRAM: 9GB, Context: 128K, License: other, LLM Explorer Score: 0.18.

  Arxiv:2009.03300   Arxiv:2407.14679   Arxiv:2408.11796   En   Endpoints compatible   Llama   Llama-3   Nemo   Nvidia   Pytorch   Region:us   Safetensors   Sharded   Tensorflow

Llama 3.1 Minitron 4B Width Base Parameters and Internals

Model Type 
text-to-text, generative
Use Cases 
Areas:
research, commercial applications
Applications:
text generation, question answering
Primary Use Cases:
Paragraph completion
Limitations:
Potential for generating toxic, biased responses
Considerations:
Security measures and ethical reviews may be needed
Supported Languages 
English (high), Multilingual (medium)
Training Details 
Data Sources:
webpages, dialogue, articles
Data Volume:
94 billion tokens
Methodology:
Knowledge distillation
Context Length:
8000
Training Time:
July 29, 2024 and Aug 3, 2024
Hardware Used:
NVIDIA A100
Model Architecture:
Llama-3.1 core, Transformer Decoder
Responsible Ai Considerations 
Transparency:
Some known biases
Accountability:
User responsibility
Mitigation Strategies:
Ensure requirements for use cases are met, unexpected misuse addressed
Input Output 
Input Format:
String
Accepted Modalities:
Text
Output Format:
String
Performance Tips:
Use text prompts of 8000 characters or less for best results
Release Notes 
Version:
4B
Date:
2024-08-03
Notes:
Final deployment-ready version.
LLM NameLlama 3.1 Minitron 4B Width Base
Repository πŸ€—https://huggingface.co/nvidia/Llama-3.1-Minitron-4B-Width-Base 
Model Size4b
Required VRAM9 GB
Updated2026-07-17
Maintainernvidia
Model Typellama
Model Files  5.0 GB: 1-of-2   4.0 GB: 2-of-2
Supported Languagesen
Model ArchitectureLlamaForCausalLM
Licenseother
Context Length131072
Model Max Length131072
Transformers Version4.45.0.dev0
Tokenizer ClassPreTrainedTokenizerFast
Vocabulary Size128256
Torch Data Typebfloat16

Best Alternatives to Llama 3.1 Minitron 4B Width Base

Best Alternatives
Context / RAM
Downloads
Likes
4Bcpt256K / 8.8 GB50
HoldMy4BKTO256K / 8.8 GB50
Xgen Small 4B Instruct R256K / 17.7 GB544
Xgen Small 4B Base R256K / 17.7 GB303
SJT 4B146K / 7.6 GB70
...lama 3.1 Nemotron Nano 4B V1.1128K / 9 GB3940116
Nemotron W 4b MagLight 0.1128K / 9.2 GB93
Loxa 4B128K / 16 GB80
Nemotron W 4b Halo 0.1128K / 9.2 GB223
Aura 4B128K / 9 GB4414
Note: green Score (e.g. "73.2") means that the model is better than nvidia/Llama-3.1-Minitron-4B-Width-Base.