Llama 2 13B Instruct V0.2 is an open-source language model by dfurman. Features: 13b LLM, VRAM: 0.2GB, License: llama2, Instruction-Based, LLM Explorer Score: 0.1, Arc: 60.6, HellaSwag: 82, MMLU: 55.5, GSM8K: 9.3.
Llama 2 13B Instruct V0.2 Parameters and Internals
| Model Type | | Causal language model (clm) |
|
| Additional Notes | | Enhanced performance can be achieved by adjusting prompt formatting during use. |
|
| Supported Languages | |
| Training Details |
| Data Sources: | | https://huggingface.co/datasets/jondurbin/airoboros-2.2.1, https://huggingface.co/datasets/Open-Orca/SlimOrca, https://huggingface.co/datasets/garage-bAInd/Open-Platypus |
|
| Data Volume: | | First 20k rows from each dataset |
|
| Methodology: | | Parameter-efficient finetuning |
|
| Training Time: | | ~8 hours on 1x A100 (40 GB SXM) |
|
| Hardware Used: | |
| Model Architecture: | | Parameter-efficient finetuning |
|
|
| Input Output |
| Input Format: | | Supports chat template with alternating user/assistant roles |
|
| Accepted Modalities: | |
| Output Format: | |
| Performance Tips: | | Apply prompt formatting for improved results. |
|
|
Best Alternatives to Llama 2 13B Instruct V0.2
Note: green Score (e.g. "73.2") means that the model is better than dfurman/Llama-2-13B-Instruct-v0.2.