Nemotron 3.5 Content Safety Model Model Developer: NVIDIA Corporation Model Dates: June 2, 2026 Model Overview The Nemotron 3.5 Content Safety model is a small language model (SLM) that uses Google's Gemma 3 4B it as the base and is fine tuned by NVIDIA on multimodal, multilingual, and reasoning oriented content safety datasets. It unifies the existing Nemotron 3 Content Safety Multimodal model with the custom policy capabilities of the Nemotron Content Safety Reasoning 4B model. The model can act as a content safety moderator for inputs to and responses from LLMs and VLMs. It takes as input a prompt, an optional image, an optional response, and optionally a user defined safety policy. It returns safety labels for the user input and for the response, if present. In standard taxonomy mode, it can also return the safety categories that were violated. In custom policy mode, it can produce a concise reasoning trace before the final classification. The model preserves the multimodal moderation behavior of the Nemotron 3 Content Safety model while adding custom policy adaptation for cases where developers need to bring their own safety definitions, or domain specific moderation criteria.…
We use cookies for essential functionality and analytics. You can accept or reject analytics cookies.Cookie policy