Model Overview Description: The NVIDIA Phi 4 multimodal instruct FP4 model is the quantized version of Microsoft’s Phi 4 multimodal instruct model, which is a multimodal foundation model that uses an optimized transformer architecture. For more information, please check here. The NVIDIA Phi 4 multimodal instruct FP4 model is quantized with TensorRT Model Optimizer. This model is ready for commercial/non commercial use. Third Party Community Consideration This model is not owned or developed by NVIDIA. This model has been developed and built to a third party’s requirements for this application and use case; see link to Non NVIDIA (Phi 4 multimodal instruct) Model Card. License/Terms of Use: Use of this model is governed by nvidia open model license ADDITIONAL INFORMATION: MIT License. Deployment Geography: Global, except in European Union Use Case: Developers looking to take off the shelf pre quantized models for deployment in AI Agent systems, chatbots, RAG systems, and other AI powered applications. Release Date: Huggingface 09/15/2025 via https://huggingface.co/nvidia/Phi 4 multimodal instruct FP4 Model Architecture: Architecture Type: Transformers Network Architecture: Phi4MMFor…
We use cookies for essential functionality and analytics. You can accept or reject analytics cookies.Cookie policy