Gemma 4 QAT models quantized to NVFP4 with GPTQ+iMatrix and FP8-calibrated KV cache using multilingual and tool-use data.
-
yasu-oh/gemma-4-31B-it-qat-NVFP4
Image-Text-to-Text • 18B • Updated • 1.22k • 1 -
yasu-oh/gemma-4-26B-A4B-it-qat-NVFP4
Image-Text-to-Text • 26B • Updated • 79 -
yasu-oh/gemma-4-12B-it-qat-NVFP4
Image-Text-to-Text • 12B • Updated • 238 -
yasu-oh/gemma-4-E4B-it-qat-NVFP4
Image-Text-to-Text • 8B • Updated • 74