Qwen3.5 9B AWQ Base model: Qwen/Qwen3.5 9B This repo quantizes the model using data free quantization technique. 【Dependencies / Installation】 As of 2026 02 25 , make sure your system has cuda12.8 installed. Then, create a fresh Python environment (e.g. python3.12 venv) and run: vLLM Official Guide 【vLLM Startup Command】 【Logs】 【Model Files】 File Size Last Updated 12GiB 2026 03 02 【Model Download】 【Overview】 Qwen3.5 9B [!Note] This repository contains model weights and configuration files for the post trained model in the Hugging Face Transformers format. These artifacts are compatible with Hugging Face Transformers, vLLM, SGLang, KTransformers, etc. Over recent months, we have intensified our focus on developing foundation models that deliver exceptional utility and performance. Qwen3.5 represents a significant leap forward, integrating breakthroughs in multimodal learning, architectural efficiency, reinforcement learning scale, and global accessibility to empower developers and enterprises with unprecedented capability and efficiency. Qwen3.5 Highlights Qwen3.5 features the following enhancement: Unified Vision Language Foundation : Early fusion training on multimodal tokens achi…
We use cookies for essential functionality and analytics. You can accept or reject analytics cookies.Cookie policy