Alibaba released a preview of its next flagship AI model on Monday, pushing the Qwen series further into the frontier race. Qwen 3.6-Max-Preview is the most powerful model the company has shipped so far, topping six major coding benchmarks and posting meaningful gains in world knowledge and instruction following over its predecessor, Qwen 3.6-Plus.
Introducing Qwen3.6-Max-Preview, an early preview of our next flagship model
Highlights:️ Improved agentic coding capability over Qwen3.6-Plus Stronger world knowledge and instruction following Improved real-world agent and knowledge reliability performance
This is also a shift in Alibaba’s business model, as it was known to provide powerful open-source models by default. The lower end models are still open source.
In terms of gains over Qwen 3.6-Plus, benchmarks for agentic skills put it on top of the family and other models like Calude 4.5 or GLM 5.1. The model also improved in knowledge benchmarks, with SuperGPQA (advanced reasoning) increasing by 2.3% and QwenChineseBench (Chinese language performance) by 5.3%. Instruction-following ability, measured by ToolcallFormatIFBench, put it on top of the rankings, beating Claude.
That approach is designed to cut compute costs without sacrificing output quality. Combined with Monday's Max release, the Qwen 3.6 lineup now spans Max-Preview at the top of the family, the Qwen Plus variance for balanced workloads, Flash for speed-first tasks, and 35B-A3B for those running locally.
The Max-Preview also ships with preserve_thinking, a feature that carries reasoning traces across multi-turn conversations. Alibaba specifically recommends it for agentic tasks where context continuity matters. For developers running autonomous agents or long-running code generation workflows, that is a meaningful addition.




















