LLM EXPLORER 60,828 MODELS INDEXED

Qwen1.5 MoE A2.7B Chat GPTQ Int4 by Qwen

By Qwen · 2155 downloads

Qwen1.5 MoE A2.7B Chat GPTQ Int4 is an open-source language model by Qwen. Features: 14.3b LLM, VRAM: 8.4GB, Context: 32K, License: other, MoE, Quantized, LLM Explorer Score: 0.13.

  4-bit   Chat   Conversational   En   Endpoints compatible   Gptq   Moe   Quantized   Qwen2 moe   Region:us   Safetensors   Sharded   Tensorflow

Qwen1.5 MoE A2.7B Chat GPTQ Int4 Parameters and Internals

LLM NameQwen1.5 MoE A2.7B Chat GPTQ Int4
Repository πŸ€—https://huggingface.co/Qwen/Qwen1.5-MoE-A2.7B-Chat-GPTQ-Int4 
Model Size14.3b
Required VRAM8.4 GB
Updated2026-07-19
MaintainerQwen
Model Typeqwen2_moe
Model Files  4.0 GB: 1-of-3   3.8 GB: 2-of-3   0.6 GB: 3-of-3
Supported Languagesen
GPTQ QuantizationYes
Quantization Typegptq
Model ArchitectureQwen2MoeForCausalLM
Licenseother
Context Length32768
Model Max Length32768
Transformers Version4.39.0.dev0
Tokenizer ClassQwen2Tokenizer
Padding Token<|endoftext|>
Vocabulary Size151936
Torch Data Typefloat16
Errorsreplace

Best Alternatives to Qwen1.5 MoE A2.7B Chat GPTQ Int4

Best Alternatives
Context / RAM
Downloads
Likes
Qwen1.5 MoE A2.7B Chat32K / 28.7 GB40226133
Qwen1.5 MoE A2.7B8K / 28.7 GB2230
Qwen1.5 MoE A2.7B8K / 28.7 GB209172227
...Aux Free Sft Math7k 1e 3 Gamma8K / 28.7 GB61
Qwen1.5 MoE Sft Math7k8K / 57.3 GB51
Qwen1.5 MoE A2.7B Wikihow8K / 28.7 GB114
Note: green Score (e.g. "73.2") means that the model is better than Qwen/Qwen1.5-MoE-A2.7B-Chat-GPTQ-Int4.