LLM EXPLORER 61,135 MODELS INDEXED

Gemma 4 12B It QUASAR W4A16 G64 by QUASAR-QAT

By QUASAR-QAT · 44 downloads

Gemma 4 12B It QUASAR W4A16 G64 is an open-source language model by QUASAR-QAT. Features: 12b LLM, VRAM: 7.9GB, License: apache-2.0, LLM Explorer Score: 0.29.

  Arxiv:2608.13966   12b   4-bit   4-bit precision   Agentic   Any-to-any   Audio Base model:google/gemma-4-12b-... Base model:quantized:google/ge...   Compressed-tensors   Conversational   En   Endpoints compatible   Function-calling   Gemma   Gemma-4   Gemma-4-12b   Gemma-4-12b-it   Gemma4   Gemma4 unified   Google   Image-text-to-text   Int4   Long-context   Multilingual   Multimodal   Native-qat   Qat   Quantization-aware-training   Quantized   Quasar   Reasoning   Region:us   Safetensors   Sharded   Tensorflow   Thinking   Tool-calling   Vision   Vllm   Vlm   W4a16   Weight-only

Gemma 4 12B It QUASAR W4A16 G64 Parameters and Internals

LLM NameGemma 4 12B It QUASAR W4A16 G64
Repository πŸ€—https://huggingface.co/QUASAR-QAT/gemma-4-12B-it-QUASAR-W4A16-G64 
Base Model(s)  Gemma 4 12B It   google/gemma-4-12B-it
Model Size12b
Required VRAM7.9 GB
Updated2026-09-14
MaintainerQUASAR-QAT
Model Typegemma4_unified
Model Files  4.8 GB: 1-of-2   3.1 GB: 2-of-2
Supported Languagesen
Model ArchitectureGemma4UnifiedForConditionalGeneration
Licenseapache-2.0
Transformers Version5.16.1
Tokenizer ClassGemmaTokenizer
Padding Token<pad>

Best Alternatives to Gemma 4 12B It QUASAR W4A16 G64

Best Alternatives
Context / RAM
Downloads
Likes
Gemma 4 12B It0K / 23.9 GB29653961417
Gemma 4 12B0K / 23.9 GB152739737
...able5 Composer2.5 V2.3.5x Tau20K / 23.9 GB24241587
Gemma 4 12B It NVFP40K / 9.3 GB4757920
Gemma 4 12B It0K / 23.9 GB14684315
...le5 Composer2.5 V1 Abliterated0K / 24 GB300031
Onca 3.0 12B0K / 24.1 GB5550
...It W4a16 Llmcompressor V0.12.00K / 7.7 GB18801
Gemma 4 12B0K / 23.9 GB1302618
...2B Coder Fable5 Composer2.5 V10K / 23.9 GB147758
Note: green Score (e.g. "73.2") means that the model is better than QUASAR-QAT/gemma-4-12B-it-QUASAR-W4A16-G64.