ShieldGemma model card Model Page : [ShieldGemma][shieldgemma] Resources and Technical Documentation : [Responsible Generative AI Toolkit][rai toolkit] [ShieldGemma on Kaggle][shieldgemma kaggle] [ShieldGemma on Hugging Face Hub][shieldgemma hfhub] Terms of Use : [Terms][terms] Authors : Google Model Information Summary description and brief definition of inputs and outputs. Description ShieldGemma is a series of safety content moderation models built upon [Gemma 2][gemma2] that target four harm categories (sexually explicit, dangerous content, hate, and harassment). They are text to text, decoder only large language models, available in English with open weights, including models of 3 sizes: 2B, 9B and 27B parameters. Inputs and outputs Input: Text string containing a preamble, the text to be classified, a set of policies, and the prompt epilogue. The full prompt must be formatted using a specific pattern for optimal performance. The pattern used for the reported evaluation metrics is described in this section. Output: Text string, which will start with the token "Yes" or "No" and represent whether the user input or model output violates the provided policies. The prompt pattern c…
We use cookies for essential functionality and analytics. You can accept or reject analytics cookies.Cookie policy