Qwen3 4B Instruct 2507 GPTQ Int4 Base model: Qwen/Qwen3 4B Instruct 2507 This model is quantized to 4 bit with a group size of 128. Compared to earlier quantized versions, the new quantized model demonstrates better tokens/s efficiency. This improvement comes from setting desc act=False in the quantization configuration. 【Dependencies】 【Model Download】 【Overview】 Qwen3 4B Instruct 2507 Highlights We introduce the updated version of the Qwen3 4B non thinking mode , named Qwen3 4B Instruct 2507 , featuring the following key enhancements: Significant improvements in general capabilities, including instruction following, logical reasoning, text comprehension, mathematics, science, coding and tool usage . Substantial gains in long tail knowledge coverage across multiple languages . Markedly better alignment with user preferences in subjective and open ended tasks , enabling more helpful responses and higher quality text generation. Enhanced capabilities in 256K long context understanding . Model Overview Qwen3 4B Instruct 2507 has the following features: Type: Causal Language Models Training Stage: Pretraining & Post training Number of Parameters: 4.0B Number of Paramaters (Non Embeddin…
We use cookies for essential functionality and analytics. You can accept or reject analytics cookies.Cookie policy