EXPERIMENTAL Updated AWQ quantization: done by stelterlab in INT4 GEMM with llm compressor (https://github.com/vllm project/llm compressor) from the vllm project. Still trying to get familiar with llm compressor. Original Weights by Qwen AI. Original Model Card follows: Qwen3 Coder 30B A3B Instruct Highlights Qwen3 Coder is available in multiple sizes. Today, we're excited to introduce Qwen3 Coder 30B A3B Instruct . This streamlined model maintains impressive performance and efficiency, featuring the following key enhancements: Significant Performance among open models on Agentic Coding , Agentic Browser Use , and other foundational coding tasks. Long context Capabilities with native support for 256K tokens, extendable up to 1M tokens using Yarn, optimized for repository scale understanding. Agentic Coding supporting for most platform such as Qwen Code , CLINE , featuring a specially designed function call format. Model Overview Qwen3 Coder 30B A3B Instruct has the following features: Type: Causal Language Models Training Stage: Pretraining & Post training Number of Parameters: 30.5B in total and 3.3B activated Number of Layers: 48 Number of Attention Heads (GQA): 32 for Q and 4 fo…
We use cookies for essential functionality and analytics. You can accept or reject analytics cookies.Cookie policy