Llama 2 13B is an open-source language model by GrazittiInteractive. Features: 13b LLM, VRAM: 7.4GB, Context: 4K, License: other, Quantized, LLM Explorer Score: 0.1, Arc: 59, HellaSwag: 82.3, MMLU: 55.4, GSM8K: 10.
Llama 2 13B Parameters and Internals
| Model Type | |
| Use Cases |
| Areas: | | Research, Commercial applications |
|
| Primary Use Cases: | | Assistant-like chat, Natural language generation tasks |
|
|
| Additional Notes | | The provided files are for a quantized version in 4-bit GGML format. |
|
| Training Details |
| Data Sources: | | publicly available online data |
|
| Data Volume: | | 2 trillion tokens (pretraining) |
|
| Methodology: | | Supervised fine-tuning (SFT) and reinforcement learning with human feedback (RLHF) |
|
| Context Length: | |
| Hardware Used: | | Meta's Research Super Cluster, A100-80GB GPUs |
|
| Model Architecture: | | Auto-regressive transformer architecture |
|
|
| Safety Evaluation |
| Ethical Considerations: | | Requires safety testing and tuning before deployment |
|
|
| Input Output |
| Input Format: | |
| Accepted Modalities: | |
| Output Format: | |
|
Best Alternatives to Llama 2 13B
Note: green Score (e.g. "73.2") means that the model is better than GrazittiInteractive/llama-2-13b.