Qwen3.5 397B A17B GPTQ Int4 [!Note] This repository contains int4 quantized model weights and configuration files for the post trained model in the Hugging Face Transformers format. These artifacts are compatible with Hugging Face Transformers, vLLM, SGLang, KTransformers, etc. [!Tip] For users seeking managed, scalable inference without infrastructure maintenance, the official Qwen API service is provided by Alibaba Cloud Model Studio. In particular, Qwen3.5 Plus is the hosted version corresponding to Qwen3.5 397B A17B with more production features, e.g., 1M context length by default, official built in tools, and adaptive tool use. For more information, please refer to the User Guide. Over recent months, we have intensified our focus on developing foundation models that deliver exceptional utility and performance. Qwen3.5 represents a significant leap forward, integrating breakthroughs in multimodal learning, architectural efficiency, reinforcement learning scale, and global accessibility to empower developers and enterprises with unprecedented capability and efficiency. Qwen3.5 Highlights Qwen3.5 features the following enhancement: Unified Vision Language Foundation : Early fusio…
We use cookies for essential functionality and analytics. You can accept or reject analytics cookies.Cookie policy