Moonshot AI and Alibaba release 2.8T and 2.4T parameter models, claim performance near GPT-5.6 and Claude Fable 5
Within days, Moonshot AI and Alibaba unveiled what they claim are frontier-class models. Moonshot's Kimi K3, at 2.8 trillion parameters, and Alibaba's Qwen3.8, at 2.4 trillion parameters, will both be released as open-weight models with full weights available for download.
Moonshot AI and Alibaba Release Massive Open-Weight Models
Moonshot AI and Alibaba have announced two large-scale AI models they claim can compete with OpenAI's GPT-5.6 Sol and Anthropic's Claude Fable 5, with both companies committing to open-weight releases.
Kimi K3: 2.8 Trillion Parameters
Beijing-based Moonshot AI released Kimi K3 on July 19, 2026, describing it as the world's largest open-source AI system at 2.8 trillion parameters. According to Moonshot's internal testing, K3 ranks above nearly every US model except GPT-5.6 Sol and Claude Fable 5, though it outperformed both on certain unspecified benchmarks. Full model weights will be released July 27, 2026.
Qwen3.8: 2.4 Trillion Parameters
Alibaba followed with a preview of Qwen3.8 over the weekend, claiming it is "one of the most powerful models available today" and "second only to Fable 5." The 2.4 trillion parameter model is described as "continuously evolving" and will be released as open-weight "soon," with no specific date provided.
Pricing Not Yet Disclosed
Neither company has announced pricing for API access to these models. Context window sizes, benchmark scores, and other technical specifications remain undisclosed pending full release.
Open-Weight vs. Proprietary
Both Chinese companies are emphasizing their commitment to making model weights publicly available for download and modification, contrasting with OpenAI and Anthropic's proprietary approach. Neither US company discloses parameter counts for their frontier models.
The releases follow DeepSeek's low-cost model announcement last year, which similarly claimed performance competitive with leading US systems.
US Export Controls Context
The announcements come as Washington restricts China's access to advanced chips through export controls. The US government recently forced Anthropic to pull its most capable system from certain markets over concerns about foreign competitors.
What This Means
These releases represent the most significant challenge to US AI labs' frontier position since DeepSeek. The 2.8T and 2.4T parameter counts substantially exceed most publicly known model sizes, though parameter count alone doesn't determine capability. Independent benchmarking will be critical once full weights are released July 27th and beyond. The open-weight strategy directly challenges the proprietary approach of OpenAI and Anthropic, potentially accelerating global AI development outside US control. If the performance claims hold up under independent testing, these models could shift the competitive landscape significantly, particularly for developers seeking alternatives to expensive API-based services.
Related Articles
OpenAI Ships GPT-6 Astra, But Executives Admit They Can't Fully Monitor What It's Thinking
OpenAI released GPT-6 Astra on Thursday, a model president Greg Brockman says could mark the start of AGI. But the model writes out its reasoning less often than prior versions, and OpenAI's chief scientist says monitoring AI thought processes will keep getting harder.
OpenAI Launches GPT-6 Astra, Claims SOTA Computer Use and Coding — But Independent Tests Show Mixed Gains at Higher Cost
OpenAI released GPT-6 Astra on September 3, 2026, claiming state-of-the-art computer use and coding performance alongside new alignment techniques. Independent evaluators found real but uneven gains, higher per-task costs, and reduced chain-of-thought monitorability.
OpenAI Releases GPT-6 Astra, First Model to Cross 'Critical' Cybersecurity Threshold
OpenAI has begun rolling out GPT-6 Astra, the first model to reach the company's internal 'Critical' cybersecurity threshold. Access is being phased, with companies in OpenAI's Daybreak cybersecurity program getting priority following added safeguards after a prior model containment breach.
OpenAI's GPT-6 Astra Reportedly Automates AI Engineering Tasks at Under $6 an Hour, According to Latent Space Testing
A Latent Space report describes GPT-6 Astra, a new OpenAI model the blog says can autonomously handle AI engineering tasks—training models, labeling data, deploying systems—at an estimated cost of under $6 per hour. The claims, including 97.6% on FrontierMath and 99.9% on ARC-AGI-3, come from independent blog testing rather than an official OpenAI announcement.
Comments
Loading...