WizardLM 30B GPTQ is an open-source language model by TheBloke. Features: 30b LLM, VRAM: 16.9GB, Context: 2K, License: other, Quantized, LLM Explorer Score: 0.12, Arc: 28.8, HellaSwag: 26.1, MMLU: 24.6, GSM8K: 34.4.
WizardLM 30B GPTQ Parameters and Internals
Model Type text generation, quantized model
Additional Notes The model consists of GPTQ 4bit quantized files suitable for GPU inference. Provides optimized versions like GGML for CPU and AutoGPTQ for more flexibility.
Input Output
Input Format: A chat between a curious user and an artificial intelligence assistant. The assistant gives helpful, detailed, and polite answers to the user's questions. USER: [prompt goes here] ASSISTANT:
Accepted Modalities:
Output Format:
LLM Name WizardLM 30B GPTQ Repository π€ https://huggingface.co/TheBloke/WizardLM-30B-GPTQ Model Size 30b Required VRAM 16.9 GB Updated 2026-08-02 Maintainer TheBloke Model Type llama Model Files 16.9 GB GPTQ Quantization Yes Quantization Type gptq Model Architecture LlamaForCausalLM License other Context Length 2048 Model Max Length 2048 Transformers Version 4.30.0.dev0 Tokenizer Class LlamaTokenizer Beginning of Sentence Token <s> End of Sentence Token </s> Unk Token <unk> Vocabulary Size 32001 Torch Data Type float16
Rank the WizardLM 30B GPTQ capabilities
Have you tried this model? Rate its performance β this feedback helps the ML community find the right model for their needs.
Instruction Following and Task Automation
Factuality and Completeness of Knowledge
Censorship and Alignment
Data Analysis and Insight Generation
Text Generation
Text Summarization and Feature Extraction
Code Generation
Multi-Language Support and Translation
Best Alternatives to WizardLM 30B GPTQ
Expand