LLM EXPLORER 59,516 MODELS INDEXED

V3 Gptq by NotoriousH2

By NotoriousH2 · 5 downloads

V3 Gptq is an open-source language model by NotoriousH2. Features: LLM, VRAM: 6.1GB, Context: 4K, Quantized, Merged, LLM Explorer Score: 0.12.

  Merged Model   Arxiv:1910.09700   4-bit   4bit   Endpoints compatible   Gptq   Llama   Quantized   Region:us
Model Card on HF πŸ€—: https://huggingface.co/NotoriousH2/v3_gptq 

V3 Gptq Parameters and Internals

LLM NameV3 Gptq
Repository πŸ€—https://huggingface.co/NotoriousH2/v3_gptq 
Merged ModelYes
Required VRAM6.1 GB
Updated2026-08-06
MaintainerNotoriousH2
Model Typellama
Model Files  6.1 GB
GPTQ QuantizationYes
Quantization Typegptq|4bit
Model ArchitectureLlamaForCausalLM
Context Length4096
Model Max Length4096
Transformers Version4.38.2
Tokenizer ClassLlamaTokenizer
Padding Token</s>
Vocabulary Size40960
Torch Data Typefloat16

Best Alternatives to V3 Gptq

Best Alternatives
Context / RAM
Downloads
Likes
LWM Text Chat 512K GPTQ512K / 4.3 GB12
LWM Text Chat 256K GPTQ256K / 4.3 GB41
LWM Text Chat 128K GPTQ128K / 4.3 GB71
StoryTeller10.7B GPTQ 4Bit41K / 6.6 GB100
Alpha Merged Gptq19K / 9.2 GB80
...lama 3 Lima Nsfw 16K Test GPTQ16K / 5.7 GB115
Taiwan LLaMa V1.0 4bits GPTQ4K / 7.3 GB101
MythoMax22b Falseblock GPT4K / 12 GB60
Taiwan LLaMa V1.0 4bits GPTQ4K / 7.3 GB89
Nous Hermes Llama2 8bit GPTQ4K / 13.7 GB41
Note: green Score (e.g. "73.2") means that the model is better than NotoriousH2/v3_gptq.