LLM EXPLORER 59,516 MODELS INDEXED

Gemma 4 31B It Qat Assistant Mxfp8 by mlx-community

By mlx-community · 277 downloads

Gemma 4 31B It Qat Assistant Mxfp8 is an open-source language model by mlx-community. Features: 31b LLM, VRAM: 0.5GB, License: gemma, Quantized, LLM Explorer Score: 0.26.

  8-bit Base model:google/gemma-4-31b-... Base model:quantized:google/ge...   Conversational   Draft-model   En   Gemma   Gemma-4   Gemma-4-31b   Gemma4 assistant   Mlx   Mlx-vlm   Mtp   Mxfp8   Q4   Q4 0   Quantized   Region:us   Safetensors   Speculative-decoding

Gemma 4 31B It Qat Assistant Mxfp8 Parameters and Internals

LLM NameGemma 4 31B It Qat Assistant Mxfp8
Repository πŸ€—https://huggingface.co/mlx-community/gemma-4-31B-it-qat-assistant-mxfp8 
Base Model(s)  ...Qat Q4 0 Unquantized Assistant   google/gemma-4-31B-it-qat-q4_0-unquantized-assistant
Model Size31b
Required VRAM0.5 GB
Updated2026-08-10
Maintainermlx-community
Model Typegemma4_assistant
Model Files  0.5 GB
Supported Languagesen
Quantization Typeq4|q4_0
Model ArchitectureGemma4AssistantForCausalLM
Licensegemma
Transformers Version5.10.0.dev0
Tokenizer ClassGemmaTokenizer
Padding Token<pad>

Best Alternatives to Gemma 4 31B It Qat Assistant Mxfp8

Best Alternatives
Context / RAM
Downloads
Likes
...Qat Q4 0 Unquantized Assistant0K / 0.9 GB2577822
...ma 4 31B It Qat Assistant Bf160K / 0.9 GB4951
Gemma 4 31B It Assistant0K / 0.9 GB1539849315
Gemma 4 31B It Assistant Bf160K / 0.9 GB297514
Gemma 4 31B It Assistant Fp80K / 0.8 GB1992
Note: green Score (e.g. "73.2") means that the model is better than mlx-community/gemma-4-31B-it-qat-assistant-mxfp8.