Qwen3 VL 2B Instruct FP8 This repository contains an FP8 quantized version of the Qwen3 VL 2B Instruct model. The quantization method is fine grained fp8 quantization with block size of 128, and its performance metrics are nearly identical to those of the original BF16 model. Enjoy! Meet Qwen3 VL — the most powerful vision language model in the Qwen series to date. This generation delivers comprehensive upgrades across the board: superior text understanding & generation, deeper visual perception & reasoning, extended context length, enhanced spatial and video dynamics comprehension, and stronger agent interaction capabilities. Available in Dense and MoE architectures that scale from edge to cloud, with Instruct and reasoning‑enhanced Thinking editions for flexible, on‑demand deployment. Key Enhancements: Visual Agent : Operates PC/mobile GUIs—recognizes elements, understands functions, invokes tools, completes tasks. Visual Coding Boost : Generates Draw.io/HTML/CSS/JS from images/videos. Advanced Spatial Perception : Judges object positions, viewpoints, and occlusions; provides stronger 2D grounding and enables 3D grounding for spatial reasoning and embodied AI. Long Context & Vide…
We use cookies for essential functionality and analytics. You can accept or reject analytics cookies.Cookie policy