NVIDIA Nemotron 3 Nano 30B A3B NVFP4 Model Overview Model Developer: NVIDIA Corporation Model Dates: September 2025 \ December 2025 Data Freshness: The post training data has a cutoff date of November 28, 2025\. The pre training data has a cutoff date of June 25, 2025\. Description Nemotron Nano 3 30B A3B NVFP4 is a quantized version of Nemotron Nano 3 30B A3B and is a large language model (LLM) trained from scratch by NVIDIA, and designed as a unified model for both reasoning and non reasoning tasks. It responds to user queries and tasks by first generating a reasoning trace and then concluding with a final response. The model's reasoning capabilities can be configured through a flag in the chat template. If the user prefers the model to provide its final answer without intermediate reasoning traces, it can be configured to do so, albeit with a slight decrease in accuracy for harder prompts that require reasoning. Conversely, allowing the model to generate reasoning traces first generally results in higher quality final solutions to queries and tasks. The model employs a hybrid Mixture of Experts (MoE) architecture, consisting of 23 Mamba 2 and MoE layers, along with 6 Attention l…
We use cookies for essential functionality and analytics. You can accept or reject analytics cookies.Cookie policy