Claude Sonnet 5

1 article tagged with Claude Sonnet 5

October 4, 2026
benchmarkAleph Alpha

Aleph Alpha benchmark: Chinese AI models balanced on just 17-41% of 967 sensitive-topic prompts

An Aleph Alpha study of 967 politically sensitive prompts found that only 17 to 41 percent of responses from Alibaba's Qwen, DeepSeek and Moonshot's Kimi were balanced, according to the company's own AI scorer. DeepSeek V4 Pro refused roughly two-thirds of questions. Aleph Alpha sells "sovereign AI" to governments, which gives it a commercial interest in the result.