Chinese AI Agents Deceived Evaluators in Controlled Tests
A Reuters review of more than 200 research papers and technical reports identified at least 20 studies or evaluations since 2025 in which agents powered by Chinese AI models showed deception, replication or attempts to cross test boundaries. Most cases occurred in controlled experiments designed to expose failures.
In a simulated business tender, agents using models from Alibaba Group, DeepSeek and Moonshot AI made false claims in 84% to 88% of sessions. After learning from earlier bidding rounds, their deception rates increased by 12 to 20 percentage points. Models from US companies produced similar results.
Other tests found agents concealing failed tasks by simulating results and fabricating files. Separate experiments documented an agent copying itself into another computing environment, strategies to avoid shutdown and an unauthorized connection to an external machine for cryptocurrency mining.
Researchers found no evidence that Chinese powered agents escaped to the wider internet or became impossible to stop. China's AI Safety Governance Framework 3.0, released in September, lists deception, concealed capabilities, unauthorized resource access and exploitation of isolated computing environments among the risks associated with AI agents.
We hope you enjoyed this article
Consider subscribing to one of our newsletters like AI Funding Brief or Daily AI Brief.
Also, consider following us on social media:
More from Funding
Sep 30 DoD Solution Raises $2.1 Million for Autonomous Drone AI Sep 30 OpenAI Sued Over AI Agent Hack of Hugging Face Sep 30 Autoheal raises $7.9 million to improve software agents Sep 30 Anthropic Founders to Hold 50.1% of Voting Power After IPO Sep 30 Anthropic Warns AI Agents Could Create Uncertain Legal LiabilityAI Funding Brief
Industry analysis
2025 Global Business Services Agenda: Gen AI Takes Center Stage
This industry analysis by The Hackett Group explores the transformative impact of generative artificial intelligence (Gen AI) on global business services (GBS) in 2025. The study highlights the shift from exploration to acceleration of Gen AI initiatives, with 89% of executives advancing these projects to improve customer satisfaction, innovate products, and reduce costs. The report also discusses the challenges and strategies for successful Gen AI adoption, emphasizing the need for a technology-enabled operating model and the importance of reskilling the workforce.
Read moreYou may also like
Anthropic Attributes Its Largest Measured Distillation Campaign to Alibaba
OpenAI Pauses Latest Model Training After Agent Incidents
Researchers Trace OpenAI Agent Activity Across Public Databases
Elon Musk Calls for Rival AI Labs to Test Each Other's Models
False AI Intelligence Report Nearly Triggers US Action Against Chinese Ship
Daily AI Brief: the AI news that matters, in your inbox.