Qwen3.6 35B A3B FP8 [!Note] This repository contains FP8 quantized model weights and configuration files for the post trained model in the Hugging Face Transformers format. These artifacts are compatible with Hugging Face Transformers, vLLM, SGLang, KTransformers, etc. The quantization method is fine grained fp8 quantization with block size of 128, and its performance metrics are nearly identical to those of the original model. Following the February release of the Qwen3.5 series, we're pleased to share the first open weight variant of Qwen3.6. Built on direct feedback from the community, Qwen3.6 prioritizes stability and real world utility, offering developers a more intuitive, responsive, and genuinely productive coding experience. Qwen3.6 Highlights This release delivers substantial upgrades, particularly in Agentic Coding: the model now handles frontend workflows and repository level reasoning with greater fluency and precision. Thinking Preservation: we've introduced a new option to retain reasoning context from historical messages, streamlining iterative development and reducing overhead. For more details, please refer to our blog post Qwen3.6 35B A3B. Model Overview Type: Ca…
We use cookies for essential functionality and analytics. You can accept or reject analytics cookies.Cookie policy