LLM EXPLORER 59,420 MODELS INDEXED

Gemma 4 31B IT NVFP4 by nvidia

By nvidia · 2773900 downloads

Gemma 4 31B IT NVFP4 is an open-source language model by nvidia. Features: 31b LLM, VRAM: 32.6GB, License: other, LLM Explorer Score: 0.43.

Base model:google/gemma-4-31b-... Base model:quantized:google/ge...   Conversational   Deploy:azure   Gemma-4-31b-it   Gemma4   Lighthouse   Model optimizer   Modelopt   Nvfp4   Nvidia   Quantized   Region:us   Safetensors   Sharded   Tensorflow

Gemma 4 31B IT NVFP4 Parameters and Internals

LLM NameGemma 4 31B IT NVFP4
Repository πŸ€—https://huggingface.co/nvidia/Gemma-4-31B-IT-NVFP4 
Base Model(s)  Gemma 4 31B It   google/gemma-4-31B-it
Model Size31b
Required VRAM32.6 GB
Updated2026-08-05
Maintainernvidia
Model Typegemma4
Model Files  9.9 GB: 1-of-4   10.0 GB: 2-of-4   10.0 GB: 3-of-4   2.7 GB: 4-of-4
Model ArchitectureGemma4ForConditionalGeneration
Licenseother
Transformers Version5.5.0.dev0
Tokenizer ClassGemmaTokenizer
Padding Token<pad>

Best Alternatives to Gemma 4 31B IT NVFP4

Best Alternatives
Context / RAM
Downloads
Likes
Gemma 4 31B It0K / 62.6 GB107066252780
Gemma 4 31B It0K / 62.6 GB3033922
Gemma 4 31B0K / 62.6 GB178519
Gemma 4 31B JANG 4M CRACK0K / 22.6 GB240561704
Gemma 4 31B0K / 62.6 GB694395498
Gemma 4 31B It FP8 Block0K / 33.3 GB379858244
Gemma 4 31B It NVFP40K / 23.3 GB111553957
Gemma 4 31B It NVFP40K / 24.8 GB3944031
Gemma 4 31B It Abliterated0K / 62.5 GB75135524
Gemma 4 31B It FP8 Dynamic0K / 33.3 GB22552223
Note: green Score (e.g. "73.2") means that the model is better than nvidia/Gemma-4-31B-IT-NVFP4.