Qwen3 Coder 30B A3B Instruct GPTQ Int8 Base model:Qwen3 Coder 30B A3B Instruct 【vLLM 4 GPU Single Node Launch Command】 Note: When using 4 GPUs, you must include enable expert parallel because expert tensor TP must be evenly divisible; for 2 GPUs this is not necessary. 【Dependencies】 【Model Update Date】 【Model Files】 File Size Last Updated 30GB 2025 08 01 【Model Download】 【Overview】 Qwen3 Coder 30B A3B Instruct Highlights Qwen3 Coder is available in multiple sizes. Today, we're excited to introduce Qwen3 Coder 30B A3B Instruct . This streamlined model maintains impressive performance and efficiency, featuring the following key enhancements: Significant Performance among open models on Agentic Coding , Agentic Browser Use , and other foundational coding tasks. Long context Capabilities with native support for 256K tokens, extendable up to 1M tokens using Yarn, optimized for repository scale understanding. Agentic Coding supporting for most platform such as Qwen Code , CLINE , featuring a specially designed function call format. Model Overview Qwen3 Coder 30B A3B Instruct has the following features: Type: Causal Language Models Training Stage: Pretraining & Post training Number of Par…
We use cookies for essential functionality and analytics. You can accept or reject analytics cookies.Cookie policy