LLM EXPLORER 60,783 MODELS INDEXED

Gemma 2 9B It AWQ INT4 by hugging-quants

By hugging-quants · 177232 downloads

Gemma 2 9B It AWQ INT4 is an open-source language model by hugging-quants. Features: 9b LLM, VRAM: 6.2GB, Context: 8K, License: gemma, Quantized, LLM Explorer Score: 0.22.

  4-bit   Autoawq   Awq Base model:google/gemma-2-9b-i... Base model:quantized:google/ge...   Conversational   Deploy:azure   En   Endpoints compatible   Gemma2   Google   Quantized   Region:us   Safetensors   Sharded   Tensorflow

Gemma 2 9B It AWQ INT4 Parameters and Internals

LLM NameGemma 2 9B It AWQ INT4
Repository πŸ€—https://huggingface.co/hugging-quants/gemma-2-9b-it-AWQ-INT4 
Base Model(s)  Gemma 2 9B It   google/gemma-2-9b-it
Model Size9b
Required VRAM6.2 GB
Updated2026-07-23
Maintainerhugging-quants
Model Typegemma2
Model Files  5.0 GB: 1-of-2   1.2 GB: 2-of-2
Supported Languagesen
AWQ QuantizationYes
Quantization Typeawq
Model ArchitectureGemma2ForCausalLM
Licensegemma
Context Length8192
Model Max Length8192
Transformers Version4.45.2
Tokenizer ClassGemmaTokenizer
Padding Token<pad>
Vocabulary Size256000
Torch Data Typebfloat16

Best Alternatives to Gemma 2 9B It AWQ INT4

Best Alternatives
Context / RAM
Downloads
Likes
Gemma 2 9B It AWQ 4bit8K / 8 GB00
GWQ 9B Preview8K / 18.6 GB93
Gemma 2 9B 8bit8K / 9.8 GB1789
Gemma 2 9B 4bit8K / 5.2 GB1602
Gemma 2 9B It Bnb 4bit8K / 6.1 GB1300732
SASTRI 1 9B8K / 6.1 GB70
Gemma 2 9B Bnb 4bit8K / 6.1 GB1060531
Gemma 2 9B It Finance8K / 18.6 GB50
Athena Gemma 2 2B It8K / 23.8 GB02
Gemma 2 9B It 4bit8K / 5.2 GB169601
Note: green Score (e.g. "73.2") means that the model is better than hugging-quants/gemma-2-9b-it-AWQ-INT4.