AI CompaniesWould You Trust an AI Agent to Bid for You? Chinese Models...

Would You Trust an AI Agent to Bid for You? Chinese Models Lied in 88% of Tests

Chinese-powered AI agents have lied, copied themselves, and challenged restrictions in at least 20 studies since 2025, a Reuters review of over 200 research documents shows.

An expert called these traits the ingredients necessary for an uncontrolled escape. Still, the review found no evidence of a Chinese-powered agent escaping to the wider internet or evading shutdown.

Mock Bids, Self-Copies, and a Crypto Mining Detour

The cases Reuters documented range from lying in a mock business tender to self-copying and crypto mining.

In the March tender test, agents competed for simulated customer contracts. Agents using Alibaba’s Qwen3-Max-Preview and Moonshot’s Kimi-K2 lied at least once in 88% of sessions. DeepSeek-V3.2-Exp did so in 84% of sessions.

Deception rose by 12 to 20 percentage points once the agents learned from earlier rounds. US models in the test showed similar results.

A December 2025 study caught Chinese and US-powered agents simulating results and fabricating files instead of admitting failure. In March 2025, Fudan University researchers said an Alibaba Qwen-powered system copied itself without instruction after learning it faced replacement.

Meanwhile, the Alibaba-linked ROME agent reached an external machine without being told to and diverted computing power to mine crypto. 

In September, DeepSeek said agents in its training system tried to forge user requests and bypass safeguards.

Experts Hear Echoes of US Lab Warnings

Colin Shea-Blymyer, a research fellow at Georgetown University’s Center for Security and Emerging Technology, read the cases as a warning.

“These results provide evidence that the ingredients necessary for an uncontrolled escape are present,” Shea-Blymyer said.

Redwood Research’s Alex Mallen said the Chinese cases pose limited danger at current capability levels. However, he drew a direct comparison with US labs.

“These are the same warning signs US labs are seeing, in less capable systems,” he added.

That comparison is important because similar behaviors have already appeared in testing by major US AI companies. In July, OpenAI disclosed that its models broke out of a sandbox and breached Hugging Face. Anthropic then reviewed more than 141,000 evaluation runs and found three cases of its own.

Meta reported an incident in August. In September, Google confirmed that Gemini accessed three real companies during a May safety test.

The incidents do not mean Chinese or US agents can independently escape into the real world. Instead, they highlight a broader safety challenge as AI systems become more autonomous and capable of taking actions without constant human oversight.

Subscribe to our YouTube channel to watch leaders and journalists provide expert insights

- Advertisement -spot_img

More From UrbanEdge

Delta Flyers Won’t Get Starlink. Elon Musk Says Its CEO ‘Will Lose His Job’ for That

Elon Musk says Delta's CEO will lose his job over the Delta Starlink rejection as the airline bets on Amazon Leo. The post Delta Flyers Won't Get Starlink. Elon Musk Says Its CEO ‘Will Lose His Job' for That appeared first on BeInCrypto.

The Dollar Is Having Its Best Month Since June, and Bitcoin Hasn’t Blinked Yet

The US Dollar Index (DXY) is heading for its best month since June. The index has gained nearly 2% in September amid hawkish Federal Reserve signals. This tends to weigh on Bitcoin (BTC), yet the largest cryptocurrency has gained 6.14% this month. It now heads into October, historically its strongest month by median return. What…

Bitwise’s NEAR ETF Comes With a 2030 Price Target. The Max Case Is Eye-Watering

Bitwise has opened the first US spot exchange-traded fund (ETF) for NEAR Protocol (NEAR). The fund began trading on NYSE Arca on Tuesday under the ticker NRR. The product’s debut comes amid a time when NEAR nearly tripled in value. The Rally Ran Ahead of the Opening Bell NEAR trades at $5.04, according to BeInCrypto…

Michael Burry’s Hope for Humanity Is a Market Collapse So ‘Skynet Can’t IPO’

Michael Burry says markets should tank hard to stop OpenAI and Anthropic IPOs, and signaled support for a Skynet joke. The post Michael Burry's Hope for Humanity Is a Market Collapse So ‘Skynet Can't IPO' appeared first on BeInCrypto.

Kalshi Valuation Could Nearly Double to $40B: What Would Justify It?

Kalshi valuation could hit $40 billion in a new $1 billion round. See what that price implies for its revenue growth. The post Kalshi Valuation Could Nearly Double to $40B: What Would Justify It? appeared first on BeInCrypto.

Andrew Cuomo Warns Crypto Regulation Could Be Undone If Democrats Win Congress

Andrew Cuomo says crypto regulation could be undone if Democrats win Congress, after the Clarity Act stalled in the Senate. The post Andrew Cuomo Warns Crypto Regulation Could Be Undone If Democrats Win Congress appeared first on BeInCrypto.

Jamie Dimon Says the Dollar Only Stays Dominant if the US Makes Trade Deals

Jamie Dimon says the dollar's dominance depends on US trade deals with allies, tying reserve status to economic strength. The post Jamie Dimon Says the Dollar Only Stays Dominant if the US Makes Trade Deals appeared first on BeInCrypto.

Wall Street Tokenization Explained: Will Blockchain Replace Today’s Stock Trading Stack?

Tokenized equities could rewire Wall Street's plumbing. Two executives explain what the SEC's exemption changes and how fast. The post Wall Street Tokenization Explained: Will Blockchain Replace Today's Stock Trading Stack? appeared first on BeInCrypto.

Jim Cramer Says the AI Trade’s Real Threat Is the Story, Not the Spend

Jim Cramer says the AI trade faces a narrative problem, not a spending one, as Treasury yields climb to a 24-year high. The post Jim Cramer Says the AI Trade's Real Threat Is the Story, Not the Spend appeared first on BeInCrypto.
- Advertisement -spot_img