LLM EXPLORER 59,516 MODELS INDEXED

AquilaChat2 34B 16K GPTQ by TheBloke

By TheBloke · 7 downloads

AquilaChat2 34B 16K GPTQ is an open-source language model by TheBloke. Features: 34b LLM, VRAM: 18.7GB, Context: 4K, License: other, Quantized, LLM Explorer Score: 0.09.

  4-bit   Aquila Base model:baai/aquilachat2-34... Base model:quantized:baai/aqui...   Custom code   Gptq   Quantized   Region:us   Safetensors   Sharded   Tensorflow

AquilaChat2 34B 16K GPTQ Parameters and Internals

Model Type 
aquila
Input Output 
Input Format:
A chat between a curious human and an artificial intelligence assistant. The assistant gives helpful, detailed, and polite answers to the human's questions. Human: {prompt} Assistant:
LLM NameAquilaChat2 34B 16K GPTQ
Repository πŸ€—https://huggingface.co/TheBloke/AquilaChat2-34B-16K-GPTQ 
Model NameAquilachat2 34B 16K
Model CreatorBeijing Academy of Artificial Intelligence
Base Model(s)  AquilaChat2 34B 16K   BAAI/AquilaChat2-34B-16K
Model Size34b
Required VRAM18.7 GB
Updated2026-07-25
MaintainerTheBloke
Model Typeaquila
Model Files  9.9 GB: 1-of-2   8.8 GB: 2-of-2
GPTQ QuantizationYes
Quantization Typegptq
Model ArchitectureAquilaForCausalLM
Licenseother
Context Length4096
Model Max Length4096
Transformers Version4.34.1
Tokenizer ClassGPT2Tokenizer
Vocabulary Size100008
Torch Data Typefloat16

Best Alternatives to AquilaChat2 34B 16K GPTQ

Best Alternatives
Context / RAM
Downloads
Likes
AquilaChat2 34B GPTQ4K / 18.7 GB162
AquilaChat2 34B 16K16K / 67.2 GB16826
Aquila2 34B8K / 136.4 GB31219
AquilaChat2 34B4K / 67.2 GB14847
AquilaChat2 34B 16K AWQ4K / 19.3 GB103
AquilaChat2 34B AWQ4K / 19.3 GB61
H2ogpt 16K Aquilachat2 34B4K / 67.2 GB94
Note: green Score (e.g. "73.2") means that the model is better than TheBloke/AquilaChat2-34B-16K-GPTQ.