Mpt 7B Chat Q8 is an open-source language model by Abzu. Features: 7b LLM, VRAM: 6.9GB, License: cc-by-nc-sa-4.0, Quantized, Instruction-Based, LLM Explorer Score: 0.21, ELO: 1063, Arc: 46.5, HellaSwag: 75.5, MMLU: 37.6, GSM8K: 4.1.
Mpt 7B Chat Q8 Benchmarks
nn.n% — How the model compares to the reference models: Anthropic Sonnet 3.5 ("so35"), GPT-4o ("gpt4o") or GPT-4 ("gpt4").
Mpt 7B Chat Q8 Parameters and Internals
| Model Type | | Dialogue generation, Chatbot |
|
| Additional Notes | | Trained with sequence length of 2048, using various attention and optimization techniques, able to extend sequence length up to 4096. |
|
| Training Details |
| Data Sources: | | jeffwan/sharegpt_vicuna, Hello-SimpleAI/HC3, tatsu-lab/alpaca, Anthropic/hh-rlhf, victor123/evol_instruct_70k |
|
| Methodology: | |
| Context Length: | |
| Hardware Used: | | 8 A100-80GB GPUs, 32 A100-40GB GPUs |
|
| Model Architecture: | | Modified decoder-only transformer |
|
|
Best Alternatives to Mpt 7B Chat Q8
Note: green Score (e.g. "73.2") means that the model is better than Abzu/mpt-7b-chat-q8.