LLM EXPLORER 60,645 MODELS INDEXED

Llama 3 8B Instruct Gradient 1048K by gradientai

By gradientai · 13691 downloads

Llama 3 8B Instruct Gradient 1048K is an open-source language model by gradientai. Features: 8b LLM, VRAM: 16.1GB, Context: 1024K, License: llama3, Instruction-Based, LLM Explorer Score: 0.26, Arc: 54.4, HellaSwag: 76.8, MMLU: 61.9, GSM8K: 44.4.

  Arxiv:2305.14233   Arxiv:2309.00071   Arxiv:2402.08268   Conversational   Deploy:sagemaker   Doi:10.57967/hf/3372   En   Endpoints compatible   Instruct   Llama   Llama-3   Meta   Region:us   Safetensors   Sharded   Tensorflow

Llama 3 8B Instruct Gradient 1048K Parameters and Internals

Model Type 
text-generation
Use Cases 
Areas:
commercial, research
Applications:
natural language generation
Primary Use Cases:
assistant-like chat
Limitations:
not suitable for use in languages other than English
Considerations:
developers to perform safety testing and tuning tailored to applications
Additional Notes 
Model is static and trained on an offline dataset. Future versions will focus on safety improvements.
Supported Languages 
en (primary)
Training Details 
Data Sources:
publicly available online data, SlimPajama, UltraChat
Data Volume:
1.4B tokens total
Methodology:
NTK-aware interpolation for RoPE theta optimization, progressive training on increasing context lengths, supervised fine-tuning (SFT), reinforcement learning with human feedback (RLHF)
Context Length:
1048
Hardware Used:
Crusoe Energy high performance L40S cluster
Model Architecture:
auto-regressive language model using an optimized transformer architecture
Safety Evaluation 
Methodologies:
red teaming, adversarial evaluations
Findings:
mitigations implemented to limit false refusals, CBRNE assessments
Risk Categories:
misuse, critical risks, cybersecurity, child safety
Ethical Considerations:
open approach to better, safer products, emphasis on responsible AI development
Responsible Ai Considerations 
Fairness:
openness, inclusivity, helpfulness
Transparency:
steps and best practices for safe deployment
Accountability:
developers
Mitigation Strategies:
Purple Llama solutions, Llama Guard for input-output safeguards
Input Output 
Input Format:
text only
Accepted Modalities:
text
Output Format:
text and code only
LLM NameLlama 3 8B Instruct Gradient 1048K
Repository πŸ€—https://huggingface.co/gradientai/Llama-3-8B-Instruct-Gradient-1048k 
Model Size8b
Required VRAM16.1 GB
Updated2026-07-28
Maintainergradientai
Model Typellama
Instruction-BasedYes
Model Files  5.0 GB: 1-of-4   5.0 GB: 2-of-4   4.9 GB: 3-of-4   1.2 GB: 4-of-4
Supported Languagesen
Model ArchitectureLlamaForCausalLM
Licensellama3
Context Length1048576
Model Max Length1048576
Transformers Version4.41.0.dev0
Tokenizer ClassPreTrainedTokenizerFast
Vocabulary Size128256
Torch Data Typebfloat16

Quantized Models of the Llama 3 8B Instruct Gradient 1048K

Model
Likes
Downloads
VRAM
...8B Instruct Gradient 1048K AWQ085 GB
...truct Gradient 1048K IMat GGUF66502 GB
...ent 1048K Molecule Q4 K M GGUF0464 GB
...B Instruct Gradient 1048K GGUF34453 GB
...radient 1048K AWQ 4bit Smashed175 GB

Best Alternatives to Llama 3 8B Instruct Gradient 1048K

Best Alternatives
Context / RAM
Downloads
Likes
...otron 8B UltraLong 4M Instruct4192K / 32.1 GB1135125
UltraLong Thinking4192K / 16.1 GB23
...a 3.1 8B UltraLong 4M Instruct4192K / 32.1 GB17624
...a 3.1 8B UltraLong 2M Instruct2096K / 32.1 GB8759
...otron 8B UltraLong 2M Instruct2096K / 32.1 GB12418
Cthulhu 8B V1.41048K / 16.1 GB1010
...raLong 1M Instruct Abliterated1048K / 32.1 GB49
...a 3.1 8B UltraLong 1M Instruct1048K / 32.1 GB138729
...otron 8B UltraLong 1M Instruct1048K / 32.1 GB70259
Zero Llama 3.1 8B Beta61048K / 16.1 GB71