LLM EXPLORER 55,709 MODELS INDEXED

Gemma 3 12B Pt 4bit by mlx-community

By mlx-community · 63 downloads

Gemma 3 12B Pt 4bit is an open-source language model by mlx-community. Features: 12b LLM, VRAM: 8.1GB, License: gemma, Quantized.

  4bit   Endpoints compatible   Gemma3   Image-text-to-text   Mlx   Quantized   Region:us   Safetensors   Sharded   Tensorflow

Gemma 3 12B Pt 4bit Parameters and Internals

LLM NameGemma 3 12B Pt 4bit
Repository πŸ€—https://huggingface.co/mlx-community/gemma-3-12b-pt-4bit 
Base Model(s)  google/gemma-3-12b   google/gemma-3-12b
Model Size12b
Required VRAM8.1 GB
Updated2026-08-08
Maintainermlx-community
Model Typegemma3
Model Files  5.4 GB: 1-of-2   2.7 GB: 2-of-2
Quantization Type4bit
Model ArchitectureGemma3ForConditionalGeneration
Licensegemma
Transformers Version4.50.0.dev0
Tokenizer ClassGemmaTokenizer
Padding Token<pad>
Torch Data Typebfloat16

Best Alternatives to Gemma 3 12B Pt 4bit

Best Alternatives
Context / RAM
Downloads
Likes
Translategemma 12B It 4bit0K / 6.6 GB8944
Translategemma 12B It 8bit0K / 12.5 GB5688
Gemma 3 12B It 4bit DWQ0K / 7.2 GB982
Gemma 3 12B It 4bit0K / 8.1 GB68326
Gemma 3 12B It 8bit0K / 14.4 GB2422
Gemma 3 12B Pt 8bit0K / 14.4 GB291
Fallen Gemma3 12B V1128K / 24.3 GB1824
Gemma 3 12B It0K / 24.3 GB2853479723
Safeword Casual V1 12B0K / 24.3 GB446
Gemma 3 R1 12B V10K / 24.3 GB614
Note: green Score (e.g. "73.2") means that the model is better than mlx-community/gemma-3-12b-pt-4bit.