LLM EXPLORER 56,604 MODELS INDEXED

Gemma 4 12B It Qat Mxfp8 by mlx-community

By mlx-community · 519 downloads

Gemma 4 12B It Qat Mxfp8 is an open-source language model by mlx-community. Features: 12b LLM, VRAM: 12.5GB, LLM Explorer Score: 0.27.

  8-bit   Conversational   En   Gemma4 unified   Image-text-to-text   Mlx   Region:us   Safetensors   Sharded   Tensorflow

Gemma 4 12B It Qat Mxfp8 Parameters and Internals

LLM NameGemma 4 12B It Qat Mxfp8
Repository πŸ€—https://huggingface.co/mlx-community/gemma-4-12B-it-qat-mxfp8 
Model Size12b
Required VRAM12.5 GB
Updated2026-08-08
Maintainermlx-community
Model Typegemma4_unified
Model Files  5.3 GB: 1-of-3   5.3 GB: 2-of-3   1.9 GB: 3-of-3
Supported Languagesen
Model ArchitectureGemma4UnifiedForConditionalGeneration
Transformers Version5.10.0.dev0
Tokenizer ClassGemmaTokenizer
Padding Token<pad>

Best Alternatives to Gemma 4 12B It Qat Mxfp8

Best Alternatives
Context / RAM
Downloads
Likes
Gemma 4 12B It NVFP40K / 9.3 GB4757920
Gemma 4 12B It0K / 23.9 GB14684315
...le5 Composer2.5 V1 Abliterated0K / 24 GB300031
Gemma 4 12B0K / 23.9 GB1302618
Gemma 4 12B StyleTune0K / 25.9 GB119127
Gemma 4 12B It0K / 23.9 GB20130
...jgoj Cantonese Gemma4 12B Base0K / 24.1 GB941
Gemma 4 12B It Bf160K / 23.9 GB4295
Gemma 4 12B It Mxfp40K / 5.4 GB4562
Gemma 4 12B It Untied0K / 25.9 GB620
Note: green Score (e.g. "73.2") means that the model is better than mlx-community/gemma-4-12B-it-qat-mxfp8.