Alibaba Challenges Frontier AI Leaders with Massive Qwen3.8-Max Model
4-trillion-parameter mixture-of-experts model it claims is "second only to Fable 5" among frontier AI systems, according to Computerworld.

Alibaba on Monday unveiled Qwen3.8-Max, a 2.4-trillion-parameter mixture-of-experts model it claims is "second only to Fable 5" among frontier AI systems, according to Computerworld. The launch positions the model directly against OpenAI's GPT-5.6 Sol and Anthropic's Claude Opus 4.8, with open-weight versions scheduled for release next week through Alibaba Cloud's Model Studio.
The benchmark pitch
The architecture is the familiar MoE trade-off: roughly 95 billion parameters activate per request out of a 2.4-trillion total, which Alibaba argues cuts inference cost without sacrificing capability on coding, reasoning, and multimodal tasks. Internal benchmarks published by the company compare Qwen3.8-Max against Claude Opus 4.8, Claude Fable 5, and GPT-5.6 Sol on coding evaluations including SWE-bench Pro and a proprietary suite Alibaba calls NL2Repo-Bench. Each rival was tested using its vendor's own coding harness — Claude Code for Anthropic, Codex for OpenAI — and Alibaba reports the highest published score across available configurations for each competitor.
The company also ran the model through three unsupervised, multi-day coding projects, including one it says executed for 16 days without human intervention. The enterprise framing stretches across legal compliance, financial analysis, engineering design, quantitative research, and multimodal content creation — pitched as completing entire business workflows rather than individual AI-assisted tasks.
What to pressure-test before signing
The 16-day number is the claim worth interrogating. Amit Jena, development manager for AI at Kanerika, told Computerworld the figure has been reprinted widely and "interrogated nowhere" — the unresolved questions are scope, frequency of human intervention, and whether the output survived code review. Charlie Dai, vice president and principal analyst at Forrester, framed the launch more broadly: Alibaba is closing ground on proprietary leaders, but the larger story is the maturation of open-weight models as credible alternatives for software engineering, domain customization, sovereignty, and cost-sensitive deployments.
The "open-weight" label deserves the same scrutiny. Publishing weights is a separate act from opening an API, and the practical question for any team evaluating Qwen3.8-Max is what actually arrives in Model Studio next week: a repository, a license, and a model card, or just an intention. Until those land, the open-weight commitment reads as a promise, not a deliverable.
For enterprise buyers, the near-term checklist is narrow — confirm the model card and license terms when the weights drop, benchmark the 95-billion active-parameter footprint against current Claude and GPT-5.6 inference costs on representative workloads, and replicate the 16-day unsupervised coding experiment internally before treating the headline as a capability.