Model Overview Description: The NVIDIA Qwen3.6 27B NVFP4 model is the quantized version of Alibaba's Qwen3.6 27B model, which is an auto regressive language model that uses an optimized transformer architecture. For more information, please check here. The NVIDIA Qwen3.6 27B NVFP4 model is quantized with Model Optimizer. This model is ready for commercial or non commercial use. License/Terms of Use: GOVERNING DOWNLOAD TERMS: Use of the model is governed by the Apache 2.0 Deployment Geography: Global Use Case: Developers looking to take off the shelf, pre quantized models for deployment in AI Agent systems, chatbots, RAG systems, and other AI powered applications. Release Date: Hugging Face 06/26/2026 via https://huggingface.co/nvidia/Qwen3.6 27B NVFP4 References NVIDIA Model Optimizer: https://github.com/NVIDIA/Model Optimizer Model Architecture: Architecture Type: Transformers Network Architecture: Hybrid Attention (Gated DeltaNet and Gated Attention) Number of Model Parameters: 27B Input: Input Type(s): Text, Image, Video Input Format(s): String, Red, Green, Blue (RGB), Video (MP4/WebM) Input Parameters: One Dimensional (1D), Two Dimensional (2D), Three Dimensional (3D) Other Pro…
We use cookies for essential functionality and analytics. You can accept or reject analytics cookies.Cookie policy