Qwen1.5 32B Chat AWQ is an open-source language model by Qwen. Features: 32b LLM, VRAM: 21.2GB, Context: 32K, License: other, Quantized, LLM Explorer Score: 0.3, ELO: 1203, Arc: 66, HellaSwag: 85.5, MMLU: 75, GSM8K: 7.1.
Qwen1.5 32B Chat AWQ Benchmarks
nn.n% — How the model compares to the reference models: Anthropic Sonnet 3.5 ("so35"), GPT-4o ("gpt4o") or GPT-4 ("gpt4").
Qwen1.5 32B Chat AWQ Parameters and Internals
| Model Type | |
| Additional Notes | | Stable support of 32K context length for models of all sizes. No need of 'trust_remote_code'. |
|
| Supported Languages | |
| Training Details |
| Methodology: | | supervised finetuning and direct preference optimization |
|
| Context Length: | |
| Model Architecture: | | Transformer architecture with SwiGLU activation, attention QKV bias, group query attention, mixture of sliding window attention and full attention. |
|
|
| Input Output |
| Input Format: | | tokenized input using the tokenizer |
|
| Accepted Modalities: | |
| Output Format: | |
| Performance Tips: | | If you encounter code switching or other undesirable results, use provided hyper-parameters in 'generation_config.json'. |
|
|
Best Alternatives to Qwen1.5 32B Chat AWQ