SOURCE // NEWS

Alibaba Launches Qwen3.8-Max: Challenging Claude with End-to-End Agent Power

Alibaba Launches Qwen3.8-Max: Challenging Claude with End-to-End Agent Power

Alibaba has officially launched its next-generation foundation model, Qwen3.8-Max. Boasting 2.4 trillion parameters, the model has rapidly ascended to the top tier of the authoritative LMSYS Chatbot Arena leaderboard. Its performance closely rivals Anthropic’s Claude series, even outperforming Claude Opus 5 and Fable 5 in multi-step reasoning, coding, and tool-use benchmarks designed for AI agents.

Beyond raw performance, Qwen3.8-Max offers disruptive pricing. Domestically, it costs 12 RMB per million input tokens and 36 RMB per million output tokens, with implicit cache hits priced at just 1.5 RMB. Internationally, its input and output prices are only 40% and 24% of Opus 5, respectively. Early adopters have already showcased highly cost-effective real-world projects, such as recreating a fully playable clone of Angry Birds within a week at negligible token costs.

Qwen3.8-Max shines across several key agent-oriented dimensions, beating GPT-5.6 in scientific replication, multimodal comprehension, and computer use. Its major strengths include end-to-end autonomous coding (capable of building projects from scratch in empty directories without human intervention), long-term system planning over thousands of interaction turns, rapid professional document parsing, and processing hundreds of hours of video via its multimodal agent framework.

In a rigorous real-world test, a detailed PRD (Product Requirement Document) was provided to Qwen3.8-Max to build an interactive AI news portal from scratch. Within five minutes, the model dissected requirements, chose the tech stack, wrote the code, and handled environment building. Impressively, it autonomously compiled a Debug Log of errors encountered and resolved, and ran an automated acceptance test spanning dependency checks, local execution, and feature validation without any human intervention.

[AgentUpdate Depth Analysis] The launch of Qwen3.8-Max represents a monumental shift for the AI Agent ecosystem, transitioning from simple task assistance to fully autonomous, closed-loop engineering. Its ability to self-debug, maintain detailed error logs, and conduct automated QA testing showcases a sophisticated system-level planning capacity that rivals the industry’s gold standard, Claude 5. By driving down token costs to a fraction of its proprietary competitors, Alibaba is lowering the barrier for deploying complex, multi-agent frameworks at scale. This combination of top-tier intelligence and disruptive cost-efficiency is poised to accelerate the commercialization of AI agents, shifting them from simple conversational UI layers to reliable, independent virtual colleagues capable of end-to-end software delivery.