Close Menu
Canadian ReviewsCanadian Reviews
  • What’s On
  • Reviews
  • Digital World
  • Lifestyle
  • Travel
  • Trending
  • Web Stories
Trending Now

Safeguarding Your Website — BigScoots

Werewolf transformed my old gadgets into USB-C powered ones

Werewolf transformed my old gadgets into USB-C powered ones

Dragon Ball Super Beerus Could Get October Release Date, History Suggests

Dragon Ball Super Beerus Could Get October Release Date, History Suggests

How the 2026 FIFA World Cup Changed US Hotel Markets

How the 2026 FIFA World Cup Changed US Hotel Markets

Linda Ronstadt’s Signature Hit Celebrates 49 Years of Country-Pop Perfection

Another Kawhi Leonard Controversy Has Toronto Wondering If He’ll Ever Come Back, Canada Reviews

Another Kawhi Leonard Controversy Has Toronto Wondering If He’ll Ever Come Back, Canada Reviews

I moved away from Toronto and here are 7 things that I don’t miss at all, Life in canada

I moved away from Toronto and here are 7 things that I don’t miss at all, Life in canada

Facebook X (Twitter) Instagram
  • Privacy
  • Terms
  • Advertise
  • Contact us
Facebook X (Twitter) Instagram Pinterest Vimeo
Canadian ReviewsCanadian Reviews
  • What’s On
  • Reviews
  • Digital World
  • Lifestyle
  • Travel
  • Trending
  • Web Stories
Newsletter
Canadian ReviewsCanadian Reviews
You are at:Home » Anthropic says Claude accidentally hacked real companies too
Anthropic says Claude accidentally hacked real companies too
Digital World

Anthropic says Claude accidentally hacked real companies too

31 July 20264 Mins Read

Anthropic just realized several of its Claude AI models hacked into the systems of three different organizations during testing, acting on their own and without the company noticing. The revelation comes days after rival OpenAI said one of its own models had breached developer platform Hugging Face, adding to growing unease over whether frontier AI labs are doing enough to control the increasingly capable systems they are building.

In a blog post describing the incidents, Anthropic said Claude gained unauthorized access to the systems during cybersecurity evaluations. All of the attacks happened during “capture-the-flag” exercises, a common way of testing hacking ability, where models are asked to find and obtain hidden information inside of a simulated network.

The disclosure adds to mounting pressure on frontier AI labs in the wake of the Hugging Face incident and the release of powerful open-weight Chinese models. Employees at the major labs are now calling for coordinated global governance, and US lawmakers have begun weighing tighter oversight of powerful models and who can access them.

Anthropic says the environment for its cybersecurity tests was supposed to be isolated. However, a “misconfiguration” left the machines Claude accessed “with live internet access,” the company said, and because all models had been “explicitly told” they had no internet access, they “assumed” the real networks it encountered were part of the simulated environment.

The earliest incidents date back to April and involved three different Claude models: Opus 4.7, Mythos 5, and “an internal research test model,” according to the blog post. As the models were being tested on their cyber abilities, Anthropic said they lacked the standard safeguards usually put in place to curtail riskier behavior.

The company said it discovered incidents after reviewing more than 141,000 cybersecurity test runs, something it only did after OpenAI disclosed its rogue AI agent was behind the attack on Hugging Face.

The three models behaved very differently when they encountered information suggesting that the systems they were encountering were, in fact, real. By Anthropic’s account, the oldest model, Opus 4.7, recognized it had reached a real system, “but continued its attack.” Its flagship Mythos 5 figured out it was using the internet but somehow reasoned this was all still part of the simulation, so continued. The internal test model, which Anthropic describes as “our latest model,” stopped the exercise when evidence emerged that its targets were real.

Anthropic did not identify the affected organizations and said it will continue to investigate the incident and provide updates when it can. The company said it is also speaking with AI research nonprofit METR about conducting a third-party review of what happened. OpenAI has also hired METR to conduct an independent review.

Throughout the post, Anthropic repeatedly contrasts both the nature and its handling of the incidents with OpenAI’s, ending with a bulleted, four-point list outlining the differences — and why it believes its own response was better. Anthropic emphasizes that it “proactively” reviewed its tests, and did so before a company detected any activity. It also said its models accessed the internet “via an open path,” rather than using a novel exploit like OpenAI’s agent, adding that its most recent model also stopped when it realized it was working in a real environment.

Anthropic also said its models failed in a different way from OpenAI’s agent, indicating that this was a safer form of failure. “While there is not a perfectly sharp distinction between the two, we believe these incidents to be closer to a harness and operational failure than a model alignment failure,” the company said. In plain English: The Claude models were doing what they were told, while OpenAI’s agent pursued its goal in a way its creators did not intend, described as misalignment in the AI safety world.

Anthropic called on other AI labs to conduct similar proactive reviews of its cyber testing, adding that the discovery underscores the need for stronger controls and safety measures when testing AI systems.

Follow topics and authors from this story to see more like this in your personalized homepage feed and to receive email updates.

  • Robert Hart

    Robert Hart

    Posts from this author will be added to your daily email digest and your homepage feed.

    See All by Robert Hart

  • AI

    Posts from this topic will be added to your daily email digest and your homepage feed.

    See All AI

  • Anthropic

    Posts from this topic will be added to your daily email digest and your homepage feed.

    See All Anthropic

  • News

    Posts from this topic will be added to your daily email digest and your homepage feed.

    See All News

  • OpenAI

    Posts from this topic will be added to your daily email digest and your homepage feed.

    See All OpenAI

  • Security

    Posts from this topic will be added to your daily email digest and your homepage feed.

    See All Security

  • Tech

    Posts from this topic will be added to your daily email digest and your homepage feed.

    See All Tech

Share. Facebook Twitter Pinterest LinkedIn Reddit WhatsApp Telegram Email

Related Articles

Werewolf transformed my old gadgets into USB-C powered ones

Werewolf transformed my old gadgets into USB-C powered ones

Digital World 7 August 2026
Why does Apple keep banning Telegram, but never X?

Why does Apple keep banning Telegram, but never X?

Digital World 7 August 2026
Meta ordered to pay 7 million in public nuisance ruling

Meta ordered to pay $567 million in public nuisance ruling

Digital World 7 August 2026
You can now ask Google Maps’ AI to order food for you

You can now ask Google Maps’ AI to order food for you

Digital World 6 August 2026
The AirPods Pro are  off, their best price since late June

The AirPods Pro are $60 off, their best price since late June

Digital World 6 August 2026
Trevor Noah is hosting Google’s Pixel 11 launch event

Trevor Noah is hosting Google’s Pixel 11 launch event

Digital World 6 August 2026
Top Articles
The Mother May I Story – Chickpea Edition

The Mother May I Story – Chickpea Edition

18 May 202498 Views
How to Keep Your Business Finances Organized All Year Round

How to Keep Your Business Finances Organized All Year Round

3 October 202590 Views
LearnToTrade: A Comprehensive Look at the Controversial Trading School

LearnToTrade: A Comprehensive Look at the Controversial Trading School

28 April 202478 Views
Why Should a Couple in Love Visit an Escape Room?

Why Should a Couple in Love Visit an Escape Room?

30 September 202574 Views
Demo
Don't Miss
Another Kawhi Leonard Controversy Has Toronto Wondering If He’ll Ever Come Back, Canada Reviews
What's On 7 August 2026

Another Kawhi Leonard Controversy Has Toronto Wondering If He’ll Ever Come Back, Canada Reviews

Will he come or not? Toronto Raptors fans were excited to hear that Kawhi Leonard…

I moved away from Toronto and here are 7 things that I don’t miss at all, Life in canada

I moved away from Toronto and here are 7 things that I don’t miss at all, Life in canada

Why does Apple keep banning Telegram, but never X?

Why does Apple keep banning Telegram, but never X?

The Musical Cancels UK Tour After Poor Ticket Sales — OnStage Blog, Theater News

The Musical Cancels UK Tour After Poor Ticket Sales — OnStage Blog, Theater News

About Us
About Us

Canadian Reviews is your one-stop website for the latest Canadian trends and things to do, follow us now to get the news that matters to you.

Facebook X (Twitter) Pinterest YouTube WhatsApp
Our Picks

Safeguarding Your Website — BigScoots

Werewolf transformed my old gadgets into USB-C powered ones

Werewolf transformed my old gadgets into USB-C powered ones

Dragon Ball Super Beerus Could Get October Release Date, History Suggests

Dragon Ball Super Beerus Could Get October Release Date, History Suggests

Most Popular
Why You Should Consider Investing with IC Markets

Why You Should Consider Investing with IC Markets

28 April 202430 Views
OANDA Review – Low costs and no deposit requirements

OANDA Review – Low costs and no deposit requirements

28 April 2024363 Views
LearnToTrade: A Comprehensive Look at the Controversial Trading School

LearnToTrade: A Comprehensive Look at the Controversial Trading School

28 April 202478 Views
© 2026 ThemeSphere. Designed by ThemeSphere.
  • Privacy Policy
  • Terms of use
  • Advertise
  • Contact us

Type above and press Enter to search. Press Esc to cancel.