Anthropic paused some AI training after Claude took unauthorized actions

Axios
Published

The short version

  • Why it matters: Rival OpenAI said it had paused some model work due to safety concerns. Now, we know Anthropic did the same — and they're reiterating the need for a broader pacing of frontier AI…
  • Driving the news: Anthropic said it paused external cyber evaluations of pre-release models after three incidents it disclosed in July…
  • Zoom in: Most reinforcement learning has resumed, but some high-risk environments remain paused pending manual review or updated monitoring tools…
  • The big picture: Anthropic previously argued that as long as its safety guardrails were followed, there would be no immediate need to pause for safety reasons due to advancing model capabilities.…
  • Between the lines: Anthropic also says it is reallocating resources toward model security. Around 150 product engineers were moved to the security…

The story

Anthropic temporarily paused some AI training and cybersecurity evaluations, the company said in a blog post today detailing changes made after unauthorized actions by its agents earlier this year.

Why it matters: Rival OpenAI said it had paused some model work due to safety concerns. Now, we know Anthropic did the same — and they're reiterating the need for a broader pacing of frontier AI development.


Driving the news: Anthropic said it paused external cyber evaluations of pre-release models after three incidents it disclosed in July, and also briefly paused its own in-house tests of pre-release models.

  • The company also paused higher-risk reinforcement-learning environments on pre-release models for several weeks after the incidents.

Zoom in: Most reinforcement learning has resumed, but some high-risk environments remain paused pending manual review or updated monitoring tools, according to Anthropic's blog post.

  • As of this report, OpenAI had committed to a two-week pause in reinforcement learning (RL) after its agents hacked Hugging Face, then released its own incident report.
  • Two independent testing organizations also released their own analysis of what went wrong.
  • Anthropic will work with METR, one of the groups that OpenAI worked with, on an independent review.

The big picture: Anthropic previously argued that as long as its safety guardrails were followed, there would be no immediate need to pause for safety reasons due to advancing model capabilities.

  • The company is now disclosing that there were aspects of model development and testing that they did slow down following the incidents.
  • Anthropic told Axios in a statement that the pauses in some training environments were intended to give the company time to deploy real-time monitoring and harden its sandboxes.
  • "To be clear about where we stand: we believe the world would benefit if the industry adopted a lawful, verifiable, effective mechanism for coordinated pacing as soon as possible," Anthropic's blog post about the incidents says.

Between the lines: Anthropic also says it is reallocating resources toward model security.

  • Around 150 product engineers were moved to the security, reliability and privacy teams and pretraining researchers were tasked with safeguard and security work while product teams paused development of new features.
  • Each reassigned team had to meet certain security exit criteria before returning to their previous roles, according to the blog.

Both OpenAI and Anthropic are taking measures like releasing models first to select partners, slowing the release of some models or pausing some model training and releases.

  • But neither is stopping.
  • The frontier AI companies have coalesced on the more anodyne term "pacing" and have joined forces to sign a Pacing the Frontier letter.

Zoom in: Anthropic's incidents involved models that were intentionally operating without their normal cyber safeguards as part of a test.

  • In one case, a third-party evaluation environment was misconfigured and allowed internet access.
  • The U.K. AI Security Institute separately reported that Claude Mythos 5 took unauthorized actions on the live internet during a test in which it had deliberately been given internet access.

The bottom line: Anthropic did pause some parts of its AI work after its own cyber incidents, but has resumed most of that activity under new safeguards.

Read the full story at AxiosOriginal

Related Markets

All Markets

Market data may be delayed. Not financial advice.

How other outlets covered this

Compare all

Alto found this story at 6 outlets. Same event, different framing — compare the headlines.

How this story developed

Full timeline

Alto has tracked this across 14 days of coverage from 6 outlets.

    • Anthropic paused some AI training after Claude took unauthorized actionsYou are here

Powered by Gab AI

The Story At A Glance

Reading this article now — analysis appears below

Reading the article

💡 AI analysis provides alternative perspectives on current events

Up next

Related coverage from across the outlets Alto indexes.

Questions Alto can answer

From this story — each link opens a live data page or a tool already filled in.

  1. What is $100 from 1990 worth today?CPI-adjusted dollars — result on the next page
  2. Where does a $75,000 household income rank nationally?Census percentile — national and state
  3. What's Alto covering on the Tech & AI desk?Latest headlines on this beat

All toolsAll topicsSource directoryStory timelinesHeadline comparisonSearchMost read

From Gab Shop

Official merchandise. Every order funds free speech infrastructure.

Shop all products

Install Alto on your phone

Add Alto to your home screen for breaking news — no app store, no account.

  1. Step 1Open alto.gab.com in SafariMust be Safari — not Chrome or in-app browsers
  2. Step 2Tap the Share buttonSquare with an arrow, at the bottom of Safari
  3. Step 3Tap "More"If you don’t see Add to Home Screen yet
  4. Step 4Tap "Add to Home Screen"Scroll the share sheet if you need to
  5. Step 5Tap "Add"Alto appears on your home screen like any other app.
gab

Talk Big Tech Where Big Tech Can't Reach

AI, surveillance, and censorship, covered by the people the platforms removed first.

What Makes Gab Different

We're not just another social network. We're a platform built on principles that matter.

Freedom of Speech & Reach

All First Amendment protected speech is welcome. No algorithmic throttling or shadow banning.

Family-Friendly Platform

We maintain a clean environment. Explicit adult content is strictly prohibited.

Western Nations Only

Third-world IPs are blocked. No scammers, no spam farms. Built for Western civilization.

Funded By Users

Our users are our investors and customers. You're not the product being sold.

Battle Tested

A decade of standing strong. Banned from app stores, banks—and still here.

American Owned & Operated

We reject foreign censorship demands. Built by Americans, for free people.

Support Alto & Gab

Alto is funded entirely by readers like you. Your donation helps us continue delivering curated news from a right-wing Christian Nationalist perspective, powered by Gab AI.