LLM EXPLORER 63,407 MODELS INDEXED

Gemma 2B AWQ by google

By google · 1734 downloads

Gemma 2B AWQ is an open-source language model by google. Features: 2b LLM, VRAM: 3.1GB, Context: 8K, License: other, Quantized, LLM Explorer Score: 0.15.

  Arxiv:1705.03551   Arxiv:1804.06876   Arxiv:1804.09301   Arxiv:1809.02789   Arxiv:1811.00937   Arxiv:1904.09728   Arxiv:1905.07830   Arxiv:1905.10044   Arxiv:1907.10641   Arxiv:1911.01547   Arxiv:1911.11641   Arxiv:2009.03300   Arxiv:2009.11462   Arxiv:2101.11718   Arxiv:2107.03374   Arxiv:2108.07732   Arxiv:2109.07958   Arxiv:2110.08193   Arxiv:2110.14168   Arxiv:2203.09509   Arxiv:2206.04615   Arxiv:2304.06364   Arxiv:2312.11805   4-bit   Awq   Endpoints compatible   Gemma   Quantized   Region:us   Safetensors
Model Card on HF πŸ€—: https://huggingface.co/google/gemma-2b-AWQ 

Gemma 2B AWQ Parameters and Internals

LLM NameGemma 2B AWQ
Repository πŸ€—https://huggingface.co/google/gemma-2b-AWQ 
Model Size2b
Required VRAM3.1 GB
Updated2026-09-16
Maintainergoogle
Model Typegemma
Model Files  3.1 GB
AWQ QuantizationYes
Quantization Typeawq
Model ArchitectureGemmaForCausalLM
Licenseother
Context Length8192
Model Max Length8192
Transformers Version4.39.0.dev0
Tokenizer ClassGemmaTokenizer
Padding Token<pad>
Vocabulary Size256000
Torch Data Typefloat16

Best Alternatives to Gemma 2B AWQ

Best Alternatives
Context / RAM
Downloads
Likes
... Codegemma 2B AWQ 4bit Smashed8K / 3.1 GB2340
Codegemma 1.1 2B AWQ8K / 3.1 GB120
Gemma 1.1 2B It AWQ8K / 3.1 GB41
Gemma 2B It AWQ8K / 3.1 GB950
Gemma 2B AWQ8K / 3.1 GB140
... 2B It Hermes Function Calling8K / 5.1 GB180
Octopus V2 Gguf AWQ8K / 1.2 GB23567
Octopus V2 Gguf AWQ8K / 1.2 GB15777
Vi Gemma 2B RAG8K / 5.1 GB3613
My AwesomeFinance Model8K / 2.1 GB120