LLM EXPLORER 60,702 MODELS INDEXED

Speechless Llama2 13B GPTQ by TheBloke

By TheBloke · 7 downloads

Speechless Llama2 13B GPTQ is an open-source language model by TheBloke. Features: 13b LLM, VRAM: 7.3GB, Context: 4K, License: llama2, Quantized, Instruction-Based, LLM Explorer Score: 0.08.

  Arxiv:2307.09288   4-bit Base model:quantized:uukuguy/s... Base model:uukuguy/speechless-... Dataset:garage-baind/open-plat...   Dataset:open-orca/openorca Dataset:wizardlm/wizardlm evol...   En   Facebook   Gptq   Instruct   Llama   Llama2   Meta   Pytorch   Quantized   Region:us   Safetensors

Speechless Llama2 13B GPTQ Parameters and Internals

Model Type 
llama
LLM NameSpeechless Llama2 13B GPTQ
Repository πŸ€—https://huggingface.co/TheBloke/Speechless-Llama2-13B-GPTQ 
Model NameSpeechless Llama2 13B
Model CreatorJiangwen Su
Base Model(s)  Speechless Llama2 13B   uukuguy/speechless-llama2-13b
Model Size13b
Required VRAM7.3 GB
Updated2026-07-15
MaintainerTheBloke
Model Typellama
Instruction-BasedYes
Model Files  7.3 GB
Supported Languagesen
GPTQ QuantizationYes
Quantization Typegptq
Model ArchitectureLlamaForCausalLM
Licensellama2
Context Length4096
Model Max Length4096
Transformers Version4.32.1
Tokenizer ClassLlamaTokenizer
Beginning of Sentence Token<s>
End of Sentence Token</s>
Unk Token<unk>
Vocabulary Size32000
Torch Data Typefloat16

Best Alternatives to Speechless Llama2 13B GPTQ

Best Alternatives
Context / RAM
Downloads
Likes
NexusRaven 13B GPTQ16K / 7.3 GB357
CodeLlama 13B Instruct GPTQ16K / 7.3 GB7539
Leo Hessianai 13B Chat GPTQ8K / 7.3 GB471
...sianai 13B Chat Bilingual GPTQ8K / 7.3 GB24
...lama2 13B Orca V2 8K 3166 GPTQ8K / 7.3 GB4824
Mythalion 13B GPTQ4K / 7.3 GB200552
Swallow 13B Instruct GPTQ4K / 7.5 GB62
...2 13B Ft Instruct Es Gptq 3bit4K / 5.7 GB63
Pygmalion 2 13B GPTQ4K / 7.3 GB3541
LoKuS 13B GPTQ4K / 7.3 GB712
Note: green Score (e.g. "73.2") means that the model is better than TheBloke/Speechless-Llama2-13B-GPTQ.