LLM EXPLORER 56,604 MODELS INDEXED

Qwen1.5 7B Chat AWQ G128 Int4 Asym Bf16 Onnx Ryzen Strix by amd

By amd · 9 downloads

Qwen1.5 7B Chat AWQ G128 Int4 Asym Bf16 Onnx Ryzen Strix is an open-source language model by amd. Features: 7b LLM, Context: 32K, License: mit, Quantized, ELO: 1143, Arc: 55.9, HellaSwag: 78.6, MMLU: 61.7, GSM8K: 13.2.

  Awq Base model:quantized:qwen/qwen... Base model:qwen/qwen1.5-7b-cha...   Chat   Conversational   En   Onnx   Quantized   Qwen2   Region:us

Qwen1.5 7B Chat AWQ G128 Int4 Asym Bf16 Onnx Ryzen Strix Benchmarks

nn.n% — How the model compares to the reference models: Anthropic Sonnet 3.5 ("so35"), GPT-4o ("gpt4o") or GPT-4 ("gpt4").

Qwen1.5 7B Chat AWQ G128 Int4 Asym Bf16 Onnx Ryzen Strix Parameters and Internals

LLM NameQwen1.5 7B Chat AWQ G128 Int4 Asym Bf16 Onnx Ryzen Strix
Repository πŸ€—https://huggingface.co/amd/Qwen1.5-7B-Chat-awq-g128-int4-asym-bf16-onnx-ryzen-strix 
Base Model(s)  Qwen/Qwen1.5-7B-Chat   Qwen/Qwen1.5-7B-Chat
Model Size7b
Updated2026-08-09
Maintaineramd
Model Typeqwen2
Supported Languagesen
AWQ QuantizationYes
Quantization Typeawq
Model ArchitectureQwen2ForCausalLM
Licensemit
Context Length32768
Model Max Length32768
Transformers Version4.37.0
Tokenizer ClassQwen2Tokenizer
Padding Token<|im_end|>
Vocabulary Size151936
Torch Data Typebfloat16
Errorsreplace

Best Alternatives to Qwen1.5 7B Chat AWQ G128 Int4 Asym Bf16 Onnx Ryzen Strix

Best Alternatives
Context / RAM
Downloads
Likes
Samantha Qwen2 7B AWQ128K / 5.6 GB100
Dolphin 2.9.2 Qwen2 7B AWQ128K / 5.6 GB310
Arcee Maestro 7B Preview AWQ128K / 5.6 GB121
CodeQwen1.5 7B Chat AWQ64K / 5.3 GB2614
CodeQwen1.5 7B AWQ64K / 5.3 GB652
Qwen2.5 7B Instruct AWQ32K / 5.6 GB451167648
Qwen2.5 Coder 7B Instruct AWQ32K / 5.6 GB31658124
Qwen2 7B Instruct AWQ32K / 5.6 GB2020023
Qwen1.5 7B Chat AWQ32K / 5.9 GB26013
Qwen1.5 7B AWQ W4 G12832K / 5.9 GB60
Note: green Score (e.g. "73.2") means that the model is better than amd/Qwen1.5-7B-Chat-awq-g128-int4-asym-bf16-onnx-ryzen-strix.