Model Overview Description: Gemma 4 31B IT is an open multimodal model built by Google DeepMind that handles text and image inputs, can process video as sequences of frames, and generates text output. It is designed to deliver frontier level performance for reasoning, agentic workflows, coding, and multimodal understanding on consumer GPUs and workstations, with a 256K token context window and support for over 140 languages. The model uses a hybrid attention mechanism that interleaves local sliding window and full global attention, with unified Keys and Values in global layers and Proportional RoPE (p RoPE) to support long context performance. The NVIDIA Gemma 4 31B IT NVFP4 model is quantized with NVIDIA Model Optimizer. This model is ready for commercial/non commercial use. Third Party Community Consideration This model is not owned or developed by NVIDIA. This model has been developed and built to a third party's requirements for this application and use case; see link to Non NVIDIA Gemma 4 31B IT Model Card License and Terms of Use: Apache License 2.0 Gemma Google AI for Developers Deployment Geography: Global Use Case: Use Case: Designed for text generation, chatbots and conve…
We use cookies for essential functionality and analytics. You can accept or reject analytics cookies.Cookie policy