Qwen1.5 MoE A2.7B Chat is an open-source language model by Qwen. Features: 14.3b LLM, VRAM: 28.7GB, Context: 32K, License: other, MoE, LLM Explorer Score: 0.24, Arc: 53.7, HellaSwag: 80.5, MMLU: 61, GSM8K: 28.2.
Qwen1.5 MoE A2.7B Chat Parameters and Internals
| Model Type | |
| Additional Notes | | The model has an inference speed 1.74 times that of Qwen1.5-7B. |
|
| Supported Languages | |
| Training Details |
| Methodology: | | Supervised finetuning and direct preference optimization |
|
| Model Architecture: | | Transformer-based MoE decoder-only |
|
|
| Input Output |
| Input Format: | | Prompt structure with system and user roles for chat |
|
| Accepted Modalities: | |
| Output Format: | |
| Performance Tips: | | Use provided hyper-parameters in `generation_config.json` for better performance in certain cases. |
|
|
Best Alternatives to Qwen1.5 MoE A2.7B Chat
Note: green Score (e.g. "73.2") means that the model is better than Qwen/Qwen1.5-MoE-A2.7B-Chat.