Gemma 3 model card Model Page : Gemma [!Note] This repository corresponds to the 12B instruction tuned version of the Gemma 3 model using Quantization Aware Training (QAT). The checkpoint in this repository is unquantized, please make sure to quantize with Q4 0 with your favorite tool Thanks to QAT, the model is able to preserve similar quality as bfloat16 while significantly reducing the memory requirements to load the model. Resources and Technical Documentation : [Gemma 3 Technical Report][g3 tech report] [Responsible Generative AI Toolkit][rai toolkit] [Gemma on Kaggle][kaggle gemma] [Gemma on Vertex Model Garden][vertex mg gemma3] Terms of Use : [Terms][terms] Authors : Google DeepMind Model Information Summary description and brief definition of inputs and outputs. Description Gemma is a family of lightweight, state of the art open models from Google, built from the same research and technology used to create the Gemini models. Gemma 3 models are multimodal, handling text and image input and generating text output, with open weights for both pre trained variants and instruction tuned variants. Gemma 3 has a large, 128K context window, multilingual support in over 140 language…
We use cookies for essential functionality and analytics. You can accept or reject analytics cookies.Cookie policy