Model Overview Description: The NVIDIA DeepSeek V4 Pro NVFP4 model is the quantized version of the DeepSeek V4 Pro model, which is a Mixture of Experts (MoE) language model with 1.6 trillion total parameters and 49 billion activated parameters. For more information, please check here. The NVIDIA DeepSeek V4 Pro NVFP4 model is quantized with Model Optimizer. This model is ready for commercial/non commercial use. Third Party Community Consideration This model is not owned or developed by NVIDIA. This model has been developed and built to a third party’s requirements for this application and use case; see link to Non NVIDIA (DeepSeek V4 Pro) Model Card. License/Terms of Use: MIT Deployment Geography: Global Use Case: DeepSeek V4 is well suited for advanced reasoning, agentic AI applications, tool use scenarios, and complex problem solving in domains such as mathematics, software engineering, and enterprise AI assistants. Release Date: Huggingface 05/27/2026 via https://huggingface.co/nvidia/DeepSeek V4 Pro NVFP4 Model Architecture: Architecture Type: Transformers Network Architecture: Mixture of Experts (MoE) with Hybrid Attention (Compressed Sparse Attention + Heavily Compressed Atte…
We use cookies for essential functionality and analytics. You can accept or reject analytics cookies.Cookie policy