Model Overview Description: The NVIDIA Qwen2.5 VL 7B Instruct FP4 model is the quantized version of Alibaba's Qwen2.5 VL 7B Instruct model, which is an auto regressive language model that uses an optimized transformer architecture. For more information, please check here. The NVIDIA Qwen2.5 VL 7B Instruct FP4 model is quantized with TensorRT Model Optimizer. This model is ready for commercial/non commercial use. Third Party Community Consideration This model is not owned or developed by NVIDIA. It was developed and built to a third party’s requirements for this application and use case. See the Non NVIDIA (Qwen2.5 VL 7B Instruct) Model Card. License/Terms of Use: Use of this model is governed by nvidia open model license ADDITIONAL INFORMATION: Apache 2.0. Deployment Geography: Global, except in European Union Use Case: Developers looking to take off the shelf pre quantized models for deployment in AI Agent systems, chatbots, RAG systems, and other AI powered applications. Release Date: Huggingface 08/22/2025 via https://huggingface.co/nvidia/Qwen2.5 VL 7B Instruct FP4 Model Architecture: Architecture Type: Transformers Network Architecture: Qwen2.5 VL 7B This model was developed b…
We use cookies for essential functionality and analytics. You can accept or reject analytics cookies.Cookie policy