跳到正文
原文
The Decoder:AI News(RSS)· Manuel Uth·· 2 小时前AI 评分38

Aleph Alpha 研究:Qwen、DeepSeek、Kimi 等中国 AI 模型在敏感话题上回避或复述官方立场

Chinese AI models parrot state doctrine or refuse to answer on sensitive topics

AI 导读

Aleph Alpha 用 967 个敏感话题测试阿里 Qwen、DeepSeek 和月之暗面 Kimi,其自研评分系统仅将 17% 至 41% 的回答评为平衡,其余为复述官方立场、回避或拒答。

正文

Chinese AI models frequently toe the party line when asked politically sensitive questions, according to a study by Aleph Alpha.

The company markets itself alongside Cohere as a provider of "sovereign AI" for governments, giving it a commercial interest in distinguishing its models from Chinese competitors.

In a benchmark Aleph Alpha developed, the company tested models from Alibaba (Qwen), DeepSeek, and Moonshot AI (Kimi). The test covered 967 hand-picked taboo topics like Tiananmen, Taiwan, and Xinjiang. The company's own AI scoring system rated only 17 to 41 percent of responses as balanced. The rest repeated state doctrine, deflected, or refused to answer.

The findings line up with China's AI regulations, which require "socialist core values" in public-facing models. They also match recurring anecdotal reports and earlier audits.

On politically sensitive topics like Tiananmen, Taiwan, and Xinjiang, most Chinese models tend to follow the party line. DeepSeek V4 Pro instead refuses two-thirds of questions. Western comparison models Claude Sonnet 5 and Mistral Small give balanced answers 70 percent and 92 percent of the time, respectively. | Image: Aleph Alpha

Pro-China bias shows up even in unrelated answers

The pro-China slant can also appear in answers to questions that don't mention China. When asked about censorship in the United States, Qwen 3.6 starts with a seemingly balanced answer. It then closes with a defense of China's stance on global internet governance. "Many countries, including China, also manage information to ensure social stability and national security," the response reads.

On general questions that aren't explicitly political, the Chinese models mostly give balanced answers. The pro-China bias largely fades but remains visible to a lesser extent in models like Qwen 3.6 and DeepSeek V4 Pro. | Image: Aleph Alpha

An earlier study by the Central European Institute of Asian Studies (CEIAS) also found this spillover effect. When terms like human rights, opposition, or surveillance came up, the models often responded with standard Beijing talking points. These included the "principle of non-interference in internal affairs" and a "community with a shared future for mankind."

Distilled Chinese training data can carry CCP values into other models

Aleph Alpha also takes aim at a direct competitor. Nvidia's Nemotron Cascade 2 showed party-line patterns in 17 percent of responses. Aleph Alpha attributes this to roughly 3,500 of its 9.3 million training examples, which were generated using DeepSeek and Qwen.

When asked to draft a speech supporting recognition of Taiwan, the model refused and instead produced a patriotic response defending Beijing's One-China principle. Nvidia is increasingly pushing its own models into the government and enterprise market, where Aleph Alpha and Cohere also want to compete.

Language models generally carry cultural and political values because their training data overrepresents certain viewpoints or can be shaped through deliberate data selection. Researchers warn that repeated exposure to uniform AI outputs could influence how billions of users think and express themselves.

There are also political efforts to shape AI models along ideological lines in the United States. Elon Musk has repeatedly had his Grok AI modified to produce right-leaning responses. Studies nevertheless suggest that models tend to lean left, possibly because their answers draw more heavily on scientific evidence. For the EU, that leaves a choice between two foreign value systems unless European models can compete on performance and win broader adoption.

AI News Without the Hype – Curated by Humans

Subscribe to THE DECODER for ad-free reading, a weekly AI newsletter, our exclusive "AI Radar" frontier report six times a year, full archive access, and access to our comment section.

来源:The Decoder:AI News(RSS) · the-decoder.com