Model Overview Description: NVIDIA Nemotron Nano v2 12B VL model enables multi image reasoning and video understanding, along with strong document intelligence, visual Q&A and summarization capabilities. This model is ready for commercial use. License/Terms of Use Governing Terms: Use of this model is governed by the NVIDIA Open Model License Agreement Deployment Geography: Global Use Case: Nemotron Nano 12B V2 VL is a model for multi modal document intelligence. It would be used by individuals or businesses that need to process documents such as invoices, receipts, and manuals. The model is capable of handling multiple images of documents, up to four images at a resolution of 1k x 2k each, along with a long text prompt. The expected use is for tasks like summarization and Visual Question Answering (VQA). The model is also expected to have a significant advantage in throughput. Release Date: Build.Nvidia.com [October 28th, 2025] via nvidia/NVIDIA Nemotron Nano VL 12B V2 Hugging Face [October 28th, 2025] via nvidia/NVIDIA Nemotron Nano VL 12B V2 BF16 Hugging Face [October 28th, 2025] via nvidia/NVIDIA Nemotron Nano VL 12B V2 FP8 Hugging Face [October 28th, 2025] via nvidia/NVIDIA Ne…
We use cookies for essential functionality and analytics. You can accept or reject analytics cookies.Cookie policy