LLM EXPLORER 59,516 MODELS INDEXED

Gemma 4 31B It Qat Assistant Bf16 by mlx-community

By mlx-community · 495 downloads

Gemma 4 31B It Qat Assistant Bf16 is an open-source language model by mlx-community. Features: 31b LLM, VRAM: 0.9GB, License: gemma, Quantized, LLM Explorer Score: 0.26.

Base model:finetune:google/gem... Base model:google/gemma-4-31b-...   Bf16   Conversational   Draft-model   En   Gemma   Gemma-4   Gemma-4-31b   Gemma4 assistant   Mlx   Mlx-vlm   Mtp   Q4   Q4 0   Quantized   Region:us   Safetensors   Speculative-decoding

Gemma 4 31B It Qat Assistant Bf16 Parameters and Internals

LLM NameGemma 4 31B It Qat Assistant Bf16
Repository πŸ€—https://huggingface.co/mlx-community/gemma-4-31B-it-qat-assistant-bf16 
Base Model(s)  ...Qat Q4 0 Unquantized Assistant   google/gemma-4-31B-it-qat-q4_0-unquantized-assistant
Model Size31b
Required VRAM0.9 GB
Updated2026-08-10
Maintainermlx-community
Model Typegemma4_assistant
Model Files  0.9 GB
Supported Languagesen
Quantization Typeq4|q4_0
Model ArchitectureGemma4AssistantForCausalLM
Licensegemma
Transformers Version5.10.0.dev0
Tokenizer ClassGemmaTokenizer
Padding Token<pad>

Best Alternatives to Gemma 4 31B It Qat Assistant Bf16

Best Alternatives
Context / RAM
Downloads
Likes
...Qat Q4 0 Unquantized Assistant0K / 0.9 GB2577822
...a 4 31B It Qat Assistant Mxfp80K / 0.5 GB2771
Gemma 4 31B It Assistant0K / 0.9 GB1539849315
Gemma 4 31B It Assistant Bf160K / 0.9 GB297514
Gemma 4 31B It Assistant Fp80K / 0.8 GB1992
Note: green Score (e.g. "73.2") means that the model is better than mlx-community/gemma-4-31B-it-qat-assistant-bf16.