Model Overview Description: The NVIDIA Qwen3 Coder Next NVFP4 model is a quantized version of Qwen's Qwen3 Coder Next model, an autoregressive language model that uses an optimized Transformer architecture with Mixture of Experts (MoE). For more information, refer to the Qwen3 Coder Next model card. The NVIDIA Qwen3 Coder Next NVFP4 model was quantized using the TensorRT Model Optimizer. This model is ready for commercial/non commercial use. Third Party Community Consideration This model is not owned or developed by NVIDIA. This model has been developed and built to a third party's requirements for this application and use case; see link to Non NVIDIA (Qwen3 Coder Next) Model Card. License/Terms of Use: MIT Deployment Geography: Global Use Case: Developers looking to take off the shelf, pre quantized models for deployment in AI Agent systems, chatbots, RAG systems, and other AI powered applications. Release Date: Huggingface via https://huggingface.co/nvidia/Qwen3 Coder Next NVFP4 Model Architecture: Architecture Type: Transformers (Hybrid) Network Architecture: Qwen3NextForCausalLM Model Details: Total Parameters: 80.1B Active Parameters: 3.1B (Sparse Mixture of Experts) Expert Co…
We use cookies for essential functionality and analytics. You can accept or reject analytics cookies.Cookie policy