Qwen3 VL 32B Instruct AWQ INT4 Model Details Quantization Details Quantization Method: AWQ Bits: 4 Group Size: 32 Calibration Dataset: HuggingFaceM4/FineVision Quantization Tool: llm compressor Get Started Prerequisite Basic Usage Additional Information Changelog v1.0.0 Initial quantized release Authors Name: Ton Cao Email: ton@cyan.kiwi Qwen3 VL 32B Instruct Meet Qwen3 VL — the most powerful vision language model in the Qwen series to date. This generation delivers comprehensive upgrades across the board: superior text understanding & generation, deeper visual perception & reasoning, extended context length, enhanced spatial and video dynamics comprehension, and stronger agent interaction capabilities. Available in Dense and MoE architectures that scale from edge to cloud, with Instruct and reasoning‑enhanced Thinking editions for flexible, on‑demand deployment. Key Enhancements: Visual Agent : Operates PC/mobile GUIs—recognizes elements, understands functions, invokes tools, completes tasks. Visual Coding Boost : Generates Draw.io/HTML/CSS/JS from images/videos. Advanced Spatial Perception : Judges object positions, viewpoints, and occlusions; provides stronger 2D grounding and e…
We use cookies for essential functionality and analytics. You can accept or reject analytics cookies.Cookie policy