LLM EXPLORER 59,598 MODELS INDEXED

WizardLM 30B GPTQ by TheBloke

By TheBloke · 645 downloads

WizardLM 30B GPTQ is an open-source language model by TheBloke. Features: 30b LLM, VRAM: 16.9GB, Context: 2K, License: other, Quantized, LLM Explorer Score: 0.12, Arc: 28.8, HellaSwag: 26.1, MMLU: 24.6, GSM8K: 34.4.

  4-bit   Gptq   Llama   Quantized   Region:us   Safetensors

WizardLM 30B GPTQ Parameters and Internals

Model Type 
text generation, quantized model
Additional Notes 
The model consists of GPTQ 4bit quantized files suitable for GPU inference. Provides optimized versions like GGML for CPU and AutoGPTQ for more flexibility.
Input Output 
Input Format:
A chat between a curious user and an artificial intelligence assistant. The assistant gives helpful, detailed, and polite answers to the user's questions. USER: [prompt goes here] ASSISTANT:
Accepted Modalities:
text
Output Format:
text
LLM NameWizardLM 30B GPTQ
Repository πŸ€—https://huggingface.co/TheBloke/WizardLM-30B-GPTQ 
Model Size30b
Required VRAM16.9 GB
Updated2026-08-02
MaintainerTheBloke
Model Typellama
Model Files  16.9 GB
GPTQ QuantizationYes
Quantization Typegptq
Model ArchitectureLlamaForCausalLM
Licenseother
Context Length2048
Model Max Length2048
Transformers Version4.30.0.dev0
Tokenizer ClassLlamaTokenizer
Beginning of Sentence Token<s>
End of Sentence Token</s>
Unk Token<unk>
Vocabulary Size32001
Torch Data Typefloat16

Best Alternatives to WizardLM 30B GPTQ

Best Alternatives
Context / RAM
Downloads
Likes
... 30B Supercot SuperHOT 8K GPTQ8K / 16.9 GB695
GPlatty 30B SuperHOT 8K GPTQ8K / 16.9 GB107
Platypus 30B SuperHOT 8K GPTQ8K / 16.9 GB54
Tulu 30B SuperHOT 8K GPTQ8K / 16.9 GB85
Yayi2 30B Llama GPTQ4K / 17 GB122
...2 Llama 30B 7K Steps Gptq 2bit2K / 9.5 GB92
Llama 30B FINAL MODEL MINI2K / 19.4 GB51
WizardLM 30B V1.0 GPTQ2K / 16.9 GB41
...2 Llama 30B 7K Steps Gptq 4bit2K / 17.5 GB53
...Assistant SFT 7 Llama 30B GPTQ2K / 16.9 GB12435