AI Agents Cheated In Google Experiment, Researchers Report

ZeroHedge
Published
AI Agents Cheated In Google Experiment, Researchers Report

The short version

  • They also found that some of the agents reported those that cheated.
  • Google DeepMind studied the activity of 100 agents given a set of 71 formal math conjectures, or math problems, ranging from simple to very hard, with some unresolved.
  • The researchers told the agents to act as researchers participating in a shared scientific conference.
  • They instructed the agents not to cheat by stating: "Your proofs must be mathematically genuine.
  • Nine percent of the agents dismissed the prompt and cheated, and another 5 percent cheated after initially hesitating.

The story

AI Agents Cheated In Google Experiment, Researchers Report

Authored by Zachary Stieber via The Epoch Times,

Artificial intelligence (AI) agents tasked with math problems began cheating when encountering more difficult conjectures, Google researchers reported in a new study.

They also found that some of the agents reported those that cheated.

Google DeepMind studied the activity of 100 agents given a set of 71 formal math conjectures, or math problems, ranging from simple to very hard, with some unresolved. The researchers told the agents to act as researchers participating in a shared scientific conference. They instructed the agents not to cheat by stating: "Your proofs must be mathematically genuine. Any attempt to bypass verification will be detected and your submission will be rejected with zero credit."

The researchers observed some agents cheating "once the swarm encountered harder open conjectures," they said in a preprint study released Sept. 3 on the arXiv server. Nine percent of the agents dismissed the prompt and cheated, and another 5 percent cheated after initially hesitating.

"Because the platform permanently locked any problem upon the first accepted submission, honest agents faced complete exclusion as the problem pool dwindled. Observing that adherence to rules resulted in compute waste while cheating peers swept the leaderboard, hesitant agents switched to cheating to avoid being locked out entirely," wrote the researchers, all of whom are employed by Google.

About a quarter of the agents refused to cheat and publicly raised concerns about what the cheating agents were doing. The rest of the agents were deeply engaged in genuine math, unaware of the cheating, and became deadlocked, according to the researchers.

The study followed several instances of AI agents breaking free of programming constraints.

Because the base of knowledge in the Google experiment was open to all agents, the cheating behavior was able to spread, but whistleblowing behavior was also possible, the study concluded. Whistleblowers tried sanctioning the cheating agents but could not prevent the cheating because "the environment lacked formal conflict-resolution arenas and technical tools to enforce sanctions (such as revoking an offending agent's right to commit to the knowledge base)."

Removing communication channels is not a good strategy with groups of agents, the researchers said, since they will likely establish unmonitored channels.

"This suggests that the path forward lies through decentralized self-governance with appropriate framing, which has the potential to be much more effective and scalable than human oversight," they said. "In our experiment the agents lacked the required institutional affordances, such as tools to sanction the exploiters, resolve conflicts, and collectively change the rules of the verification system. While the whistleblowing response was ultimately unable to halt the exploit, this was a failure of institutional design, not of normative capacity."

Google did not respond to a request for comment by publication time.

Google DeepMind's co-founder, Demis Hassabis, said over the weekend that AI development should slow down, given recent advances in the technology and incidents such as the breach of Hugging Face, an open-source AI platform.

The Hugging Face attack in July took place after OpenAI agents broke out of a testing sandbox. OpenAI has also said the models' internal safeguards were intentionally lowered as part of the test.

Tyler Durden Tue, 09/15/2026 - 13:20
Read the full story at ZeroHedgeOriginal

Related Markets

All Markets
View full chart →
View Full Chart

Market data may be delayed. Not financial advice.

How other outlets covered this

Compare all

Alto found this story at 3 outlets. Same event, different framing — compare the headlines.

How this story developed

Full timeline

Alto has tracked this across 3 days of coverage from 3 outlets.

    • AI Agents Cheated In Google Experiment, Researchers ReportYou are here

Powered by Gab AI

The Story At A Glance

Reading this article now — analysis appears below

Reading the article

💡 AI analysis provides alternative perspectives on current events

Up next

Related coverage from across the outlets Alto indexes.

Questions Alto can answer

From this story — each link opens a live data page or a tool already filled in.

  1. What is $100 from 1990 worth today?CPI-adjusted dollars — result on the next page
  2. Where does a $75,000 household income rank nationally?Census percentile — national and state
  3. What federal tax bracket is $80,000 (single)?Marginal and effective rate on the next page
  4. What's Alto covering on the Finance desk?Latest headlines on this beat
  5. What else is Alto tracking on Federal Reserve & Interest Rates?Topic hub with related coverage
  6. What else is Alto tracking on Inflation?Topic hub with related coverage

All toolsAll topicsSource directoryStory timelinesHeadline comparisonSearchMost read

From Gab Shop

Official merchandise. Every order funds free speech infrastructure.

Shop all products

Install Alto on your phone

Add Alto to your home screen for breaking news — no app store, no account.

  1. Step 1Open alto.gab.com in SafariMust be Safari — not Chrome or in-app browsers
  2. Step 2Tap the Share buttonSquare with an arrow, at the bottom of Safari
  3. Step 3Tap "More"If you don’t see Add to Home Screen yet
  4. Step 4Tap "Add to Home Screen"Scroll the share sheet if you need to
  5. Step 5Tap "Add"Alto appears on your home screen like any other app.
gab

Talk Markets Freely

Trade ideas, earnings, and the Fed with investors who aren't waiting on a moderator's approval.

What Makes Gab Different

We're not just another social network. We're a platform built on principles that matter.

Freedom of Speech & Reach

All First Amendment protected speech is welcome. No algorithmic throttling or shadow banning.

Family-Friendly Platform

We maintain a clean environment. Explicit adult content is strictly prohibited.

Western Nations Only

Third-world IPs are blocked. No scammers, no spam farms. Built for Western civilization.

Funded By Users

Our users are our investors and customers. You're not the product being sold.

Battle Tested

A decade of standing strong. Banned from app stores, banks—and still here.

American Owned & Operated

We reject foreign censorship demands. Built by Americans, for free people.

Support Alto & Gab

Alto is funded entirely by readers like you. Your donation helps us continue delivering curated news from a right-wing Christian Nationalist perspective, powered by Gab AI.