Alibaba previews Qwen3.8 with 2.4 trillion parameters, claims second place without benchmark data
Alibaba unveiled Qwen3.8 at the World Artificial Intelligence Conference in Shanghai, claiming the 2.4 trillion parameter model ranks second only to Anthropic's Fable 5. The company provided no benchmark scores, model card, or independent verification to support the claim.
Alibaba previews Qwen3.8 with 2.4 trillion parameters, claims second place without benchmark data
Alibaba unveiled Qwen3.8 at the World Artificial Intelligence Conference in Shanghai, claiming the 2.4 trillion parameter model ranks second only to Anthropic's Fable 5. The company provided no benchmark scores, model card, or independent verification to support the claim.
Qwen3.8 is Alibaba's first multimodal model above one trillion parameters, handling images, video, documents, and text. According to Alibaba, the model should outperform its previous flagship Qwen3.7-Max, particularly at coding and office tasks including full-stack development and data analysis. Lead developer Shuai Bai said the model's understanding matches or beats leading proprietary systems, though no evidence was provided.
The model is available now through Alibaba's Token Plan subscription and Qoder developer tools at 10% of standard pricing during the trial period. Alibaba promises to release full open weights "soon" but has not specified a date or license terms.
Missing verification
Alibaba shipped Qwen3.8 without benchmark scores, a model card, or any independent testing. This marks a departure from the company's previous release: Qwen3.7-Max launched in May with published results across standard benchmarks.
No public leaderboard has scored Qwen3.8. On LMArena, where Fable 5 ranks first, the older Qwen3.7-Max sits well below the top positions. Without released weights, external researchers cannot verify Alibaba's claims.
Chinese model competition intensifies
The announcement came three days after Moonshot released Kimi K3, a 2.8 trillion parameter open model that topped a major coding leaderboard. Multiple Chinese labs are now shipping frontier-scale models: Moonshot has Kimi K3, Zhipu has GLM-5.2, and most are releasing weights publicly.
Both Alibaba and Moonshot are deploying models in paid products before releasing weights, converting launch momentum into customers before open access becomes available. This represents a shift for Alibaba, which has historically kept its largest models closed.
Alibaba shares rose as much as 5.4% on Monday following the announcement.
What this means
Alibaba's announcement signals escalating competition among Chinese AI labs, with multiple companies now claiming frontier-scale capabilities and promising open weights. The lack of benchmarks or verification data represents either rushed competitive pressure or a calculated marketing move. If Alibaba releases weights matching its claims, it would push the open-weight frontier significantly forward. Until then, the "second only to Fable 5" claim remains unverified marketing. The pattern of paid deployment before weight release suggests Chinese labs are developing sustainable business models around open-weight releases rather than purely research-driven distribution.
Related Articles
OpenAI Launches GPT-6 Astra, Says the Model May Already Qualify as AGI
OpenAI has released GPT-6 Astra, its most capable model yet, with benchmark scores the company says surpass GPT-5.6 Sol and Anthropic's Fable 5 models. President Greg Brockman called it a step into the 'AGI era,' though OpenAI acknowledges there's no agreed-upon threshold for that term.
Meta's Muse Spark 1.3 Claims #3 Global Ranking, Matches OpenAI's GPT-5.6-Sol on Coding Benchmarks
Meta Superintelligence Labs shipped Muse Spark 1.3, which the company claims ranks #3 globally on the Artificial Analysis Intelligence Index and matches OpenAI's GPT-5.6-Sol on coding and agentic benchmarks. The model is available now via Muse Code and Meta's API, with open weights and a follow-up model promised soon.
Meta Releases Muse Spark 1.3 Contributor, a Low-Cost Multimodal Reasoning Model With 1M Context Window
Meta has released Muse Spark 1.3 Contributor, described as the cost-efficient contributor tier of its multimodal reasoning model line. The model offers a 1 million token context window at $0.10 per 1M input tokens and $0.20 per 1M output tokens, targeting experimentation and early-stage agentic workflows.
OpenAI Releases GPT-6 Astra, First Model to Cross 'Critical' Cybersecurity Threshold
OpenAI has begun rolling out GPT-6 Astra, the first model to reach the company's internal 'Critical' cybersecurity threshold. Access is being phased, with companies in OpenAI's Daybreak cybersecurity program getting priority following added safeguards after a prior model containment breach.
Comments
Loading...