LLM EXPLORER 60,783 MODELS INDEXED

Gemma 2 9B It AWQ 4bit by TitanML

By TitanML · 0 downloads

Gemma 2 9B It AWQ 4bit is an open-source language model by TitanML. Features: 9b LLM, VRAM: 8GB, Context: 8K, Quantized.

  4-bit   4bit   Autotrain compatible   Awq   Conversational   Endpoints compatible   Gemma2   Quantized   Region:us   Safetensors   Sharded   Tensorflow

Gemma 2 9B It AWQ 4bit Parameters and Internals

LLM NameGemma 2 9B It AWQ 4bit
Repository πŸ€—https://huggingface.co/TitanML/gemma-2-9b-it-AWQ-4bit 
Model Size9b
Required VRAM8 GB
Updated2024-07-04
MaintainerTitanML
Model Typegemma2
Model Files  6.8 GB: 1-of-2   1.2 GB: 2-of-2
AWQ QuantizationYes
Quantization Typeawq|4bit
Model ArchitectureGemma2ForCausalLM
Context Length8192
Model Max Length8192
Transformers Version4.42.0
Tokenizer ClassGemmaTokenizer
Padding Token<pad>
Vocabulary Size256000
Torch Data Typefloat16

Best Alternatives to Gemma 2 9B It AWQ 4bit

Best Alternatives
Context / RAM
Downloads
Likes
Gemma 2 9B It AWQ INT48K / 6.2 GB1772329
GWQ 9B Preview8K / 18.6 GB93
Gemma 2 9B 8bit8K / 9.8 GB1789
Gemma 2 9B 4bit8K / 5.2 GB1602
Gemma 2 9B It Bnb 4bit8K / 6.1 GB1300732
SASTRI 1 9B8K / 6.1 GB70
Gemma 2 9B Bnb 4bit8K / 6.1 GB1060531
Gemma 2 9B It Finance8K / 18.6 GB50
Athena Gemma 2 2B It8K / 23.8 GB02
Gemma 2 9B It 4bit8K / 5.2 GB169601
Note: green Score (e.g. "73.2") means that the model is better than TitanML/gemma-2-9b-it-AWQ-4bit.