LLM EXPLORER 59,420 MODELS INDEXED

Qwen1.5 72B Chat GPTQ by LoneStriker

By LoneStriker · 2 downloads

Qwen1.5 72B Chat GPTQ is an open-source language model by LoneStriker. Features: 72b LLM, VRAM: 45.4GB, Context: 32K, License: other, Quantized, LLM Explorer Score: 0.24, ELO: 1233, Arc: 68.5, HellaSwag: 86.4, MMLU: 77.4, GSM8K: 20.4.

  Arxiv:2309.16609   4bit   Chat   Conversational   En   Endpoints compatible   Gptq   Quantized   Qwen2   Region:us   Safetensors

Qwen1.5 72B Chat GPTQ Benchmarks

nn.n% — How the model compares to the reference models: Anthropic Sonnet 3.5 ("so35"), GPT-4o ("gpt4o") or GPT-4 ("gpt4").

Qwen1.5 72B Chat GPTQ Parameters and Internals

Model Type 
text generation, chat model
Additional Notes 
The beta version does not include GQA and the mixture of SWA and full attention. DPO improves human preference but lowers benchmark evaluation.
Supported Languages 
en (multilingual capabilities include English)
Training Details 
Methodology:
Supervised finetuning and direct preference optimization (DPO)
Context Length:
32000
Model Architecture:
Transformer architecture with SwiGLU activation, attention QKV bias, group query attention, mixture of sliding window attention and full attention
Input Output 
Accepted Modalities:
text
Performance Tips:
Use provided hyper-parameters in 'generation_config.json' for optimal performance.
LLM NameQwen1.5 72B Chat GPTQ
Repository πŸ€—https://huggingface.co/LoneStriker/Qwen1.5-72B-Chat-GPTQ 
Base Model(s)  Qwen/Qwen1.5-72B-Chat   Qwen/Qwen1.5-72B-Chat
Model Size72b
Required VRAM45.4 GB
Updated2026-07-28
MaintainerLoneStriker
Model Typeqwen2
Model Files  45.4 GB
Supported Languagesen
GPTQ QuantizationYes
Quantization Typegptq|4bit
Model ArchitectureQwen2ForCausalLM
Licenseother
Context Length32768
Model Max Length32768
Transformers Version4.37.1
Tokenizer ClassQwen2Tokenizer
Padding Token<|endoftext|>
Vocabulary Size152064
Torch Data Typefloat16
Errorsreplace

Best Alternatives to Qwen1.5 72B Chat GPTQ

Best Alternatives
Context / RAM
Downloads
Likes
Kimi Dev 72B GPTQ 4bit128K / 45.8 GB122
Qwen2.5 72B Instruct GPTQ Int432K / 41.6 GB4839944
Qwen2.5 72B Instruct GPTQ Int832K / 77 GB189328
Qwen2 72B Instruct GPTQ Int832K / 77 GB170215
Qwen1.5 72B Chat GPTQ Int432K / 41.3 GB626737
Qwen2 72B Instruct GPTQ Int432K / 41.6 GB9633
Qwen1.5 72B Chat GPTQ Int832K / 77 GB187
Kimi Dev 72B 4bit DWQ128K / 40.9 GB446521
Kimi Dev 72B 8bit128K / 77.1 GB3612
Kimi Dev 72B 5bit128K / 50.1 GB3682