Aleph Alpha
1 article tagged with Aleph Alpha
October 4, 2026
benchmarkAleph Alpha
Aleph Alpha benchmark: Chinese AI models balanced on just 17-41% of 967 sensitive-topic prompts
An Aleph Alpha study of 967 politically sensitive prompts found that only 17 to 41 percent of responses from Alibaba's Qwen, DeepSeek and Moonshot's Kimi were balanced, according to the company's own AI scorer. DeepSeek V4 Pro refused roughly two-thirds of questions. Aleph Alpha sells "sovereign AI" to governments, which gives it a commercial interest in the result.