LLM EXPLORER 59,420 MODELS INDEXED

Gemma 2B Gptq 4bit by elysiantech

By elysiantech · 7 downloads

Gemma 2B Gptq 4bit is an open-source language model by elysiantech. Features: 2b LLM, VRAM: 2.1GB, Context: 8K, License: other, Quantized, LLM Explorer Score: 0.13.

  Arxiv:2308.07662   4-bit   4bit   En   Endpoints compatible   Gemma   Google   Gptq   Quantized   Region:us   Safetensors

Gemma 2B Gptq 4bit Parameters and Internals

Model Type 
text-generation-inference
Additional Notes 
gemma-2b-gptq-4bit is a quantitized version using the GPTQ method.
Supported Languages 
English (Proficient)
LLM NameGemma 2B Gptq 4bit
Repository πŸ€—https://huggingface.co/elysiantech/gemma-2b-gptq-4bit 
Model Size2b
Required VRAM2.1 GB
Updated2026-05-31
Maintainerelysiantech
Model Typegemma
Model Files  2.1 GB
Supported Languagesen
GPTQ QuantizationYes
Quantization Typegptq|4bit
Model ArchitectureGemmaForCausalLM
Licenseother
Context Length8192
Model Max Length8192
Transformers Version4.41.2
Tokenizer ClassGemmaTokenizer
Padding Token<pad>
Vocabulary Size256000
Torch Data Typefloat16

Best Alternatives to Gemma 2B Gptq 4bit

Best Alternatives
Context / RAM
Downloads
Likes
Gemma 1.1 2B It GPTQ8K / 3.1 GB338341
Gemma 2B Gptq 8bit8K / 3.1 GB60
Gemma 2B GPTQ8K / 2.1 GB101
CodeGemma 2B GPTQ8K / 3.1 GB21
... 2B It Hermes Function Calling8K / 5.1 GB180
Vi Gemma 2B RAG8K / 5.1 GB3613
My AwesomeFinance Model8K / 2.1 GB70
... 2.8 Gemma 2B HQQ 1bit Smashed8K / 1.3 GB50
... 2.8 Gemma 2B HQQ 2bit Smashed8K / 1.6 GB50
... 2.8 Gemma 2B HQQ 4bit Smashed8K / 2.1 GB50
Note: green Score (e.g. "73.2") means that the model is better than elysiantech/gemma-2b-gptq-4bit.