How China is preparing for the risk of AI escaping human control

Warnings from researchers at leading U.S. AI developer Anthropic that increasingly powerful models could escape human control have drawn attention in China, where policymakers have been preparing for some of the same risks.
The U.S. and China are the two major driving forces of frontier AI development and the technology’s global adoption.
Both superpowers have been at loggerheads over AI policies and industry practices, with these issues slated to feature prominently in bilateral talks later this month.
While the debate in the U.S. is focused on whether frontier AI could pose an existential threat to humanity, Chinese policymakers have generally treated AI as a powerful but governable technology whose risks can be contained through technical standards, regulation and state oversight.
“Chinese and American experts largely agree on AI risks,” said Brian Tse, founder and CEO of Concordia AI, a Beijing- and Singapore-based AI safety and governance research group, adding the difference was on “how risks are framed and prioritised”.
China has not proposed embedding independent monitors inside AI companies, as Anthropic has advocated. Its emerging regime instead relies on developer obligations, state-backed standards, security assessments and outside testing.
That approach is also shaped by a major difference in the two countries’ AI industries. Chinese developers have increasingly promoted open-weight models. These refer to systems whose underlying parameters can be downloaded, inspected and modified. Their leading U.S. rivals such as Anthropic and OpenAI, however, do not make these specifications publicly available. Policy issued in May by China’s cyberspace regulator, economic planner and industry ministry identifies “operational loss of control” as a security risk for AI agents, systems that can plan and carry out multi-step tasks more independently than conventional chatbots. The rules require developers to improve their ability to discover, intervene in, block and recover from improper agent behaviour.
The policy also calls on developers to guard against risks including data poisoning, algorithm manipulation and system vulnerabilities. It also says users should be informed about agents’ autonomous decisions and retain final decision-making authority.
China has also begun drafting a mandatory national standard for AI agent safety, which Concordia’s Tse said would be the world’s first of its kind. Source: Reuters

Be the first to comment

Leave a Reply

Your email address will not be published.