LLM EXPLORER 60,828 MODELS INDEXED

Google Gemma 7B 8 Bit Gptq by SotiriosKastanas

By SotiriosKastanas · 5 downloads

Google Gemma 7B 8 Bit Gptq is an open-source language model by SotiriosKastanas. Features: 7b LLM, VRAM: 9.5GB, Context: 8K, Quantized, LLM Explorer Score: 0.13.

  Arxiv:1910.09700   8-bit   Endpoints compatible   Gemma   Gptq   Quantized   Region:us   Safetensors

Google Gemma 7B 8 Bit Gptq Parameters and Internals

LLM NameGoogle Gemma 7B 8 Bit Gptq
Repository πŸ€—https://huggingface.co/SotiriosKastanas/google-gemma-7b-8-bit-gptq 
Model Size7b
Required VRAM9.5 GB
Updated2026-06-27
MaintainerSotiriosKastanas
Model Typegemma
Model Files  9.5 GB
GPTQ QuantizationYes
Quantization Typegptq
Model ArchitectureGemmaForCausalLM
Context Length8192
Model Max Length8192
Transformers Version4.41.2
Tokenizer ClassGemmaTokenizer
Padding Token<pad>
Vocabulary Size256000
Torch Data Typefloat16

Best Alternatives to Google Gemma 7B 8 Bit Gptq

Best Alternatives
Context / RAM
Downloads
Likes
Gemma 1.1 7B It GPTQ8K / 7.2 GB31
Codegemma 7B It GPTQ8K / 7.2 GB41
CodeGemma 7B GPTQ8K / 7.2 GB90
Google Gemma 7B 4 Bit Gptq8K / 5.6 GB121
Gemma 7B Instruct GPTQ 4bit8K / 5.6 GB120
Codegemma 1.1 7B It GPTQ8K / 7.2 GB41
SeaLLM 7B V2.5 4bit8K / 7.2 GB62
Gemma 7B GPTQ8K / 7.2 GB890
Gemma 7B It GPTQ8K / 7.2 GB50
...t Cleaner Gemma 32k Merged 16b31K / 17.1 GB70
Note: green Score (e.g. "73.2") means that the model is better than SotiriosKastanas/google-gemma-7b-8-bit-gptq.