LLM EXPLORER 59,598 MODELS INDEXED

Tulu 30B SuperHOT 8K GPTQ by TheBloke

By TheBloke · 8 downloads

Tulu 30B SuperHOT 8K GPTQ is an open-source language model by TheBloke. Features: 30b LLM, VRAM: 16.9GB, Context: 8K, License: other, Quantized, LLM Explorer Score: 0.07.

  4-bit   Custom code   Ext 8k   Gptq   Llama   Quantized   Region:us   Safetensors

Tulu 30B SuperHOT 8K GPTQ Parameters and Internals

Model Type 
GPTQ 4bit
Additional Notes 
An experimental GPTQ offering up to 8K context size; requires specific loaders and settings for optimal use.
Input Output 
Performance Tips:
Using the full 8K context on a 30B model will exceed 24GB VRAM.
LLM NameTulu 30B SuperHOT 8K GPTQ
Repository πŸ€—https://huggingface.co/TheBloke/Tulu-30B-SuperHOT-8K-GPTQ 
Model Size30b
Required VRAM16.9 GB
Updated2026-08-01
MaintainerTheBloke
Model Typellama
Model Files  16.9 GB
GPTQ QuantizationYes
Context Length8k
Quantization Typegptq
Model ArchitectureLlamaForCausalLM
Licenseother
Context Length8192
Model Max Length8192
Transformers Version4.30.0.dev0
Tokenizer ClassLlamaTokenizer
Vocabulary Size32001
Torch Data Typefloat16

Best Alternatives to Tulu 30B SuperHOT 8K GPTQ

Best Alternatives
Context / RAM
Downloads
Likes
... 30B Supercot SuperHOT 8K GPTQ8K / 16.9 GB695
GPlatty 30B SuperHOT 8K GPTQ8K / 16.9 GB107
Platypus 30B SuperHOT 8K GPTQ8K / 16.9 GB54
Yayi2 30B Llama GPTQ4K / 17 GB122
WizardLM 30B GPTQ2K / 16.9 GB64518
...2 Llama 30B 7K Steps Gptq 2bit2K / 9.5 GB92
Llama 30B FINAL MODEL MINI2K / 19.4 GB51
WizardLM 30B V1.0 GPTQ2K / 16.9 GB41
...2 Llama 30B 7K Steps Gptq 4bit2K / 17.5 GB53
...Assistant SFT 7 Llama 30B GPTQ2K / 16.9 GB12435
Note: green Score (e.g. "73.2") means that the model is better than TheBloke/Tulu-30B-SuperHOT-8K-GPTQ.