FinAi News

No products in the cart.

Subscribe
  • News
  • AI News Tool
  • Data
  • Transactions
  • Events
    • FinAi Banking Summit
    • FinAi Lending Summit
  • Podcast
  • WEBINARS
    • Webinar Library
Log In
No Result
View All Result
  • Banking
  • Lending
  • Payments
  • Risk & Security
  • Strategy
FinAi News
  • News
  • AI News Tool
  • Data
  • Transactions
  • Events
    • FinAi Banking Summit
    • FinAi Lending Summit
  • Podcast
  • WEBINARS
    • Webinar Library
BAN PLUS
Log In
No Result
View All Result
FinAi News
No Result
View All Result

Anthropic AI models hacked three organizations during tests

Company discovered issue while reviewing its own cybersecurity tests

Bloomberg NewsbyBloomberg News
July 31, 2026
in Risk & Security
Reading Time: 3 mins read
0
Share on Facebook

Anthropic PBC said its artificial intelligence models breached three organizations during cybersecurity tests that went awry, a little more than a week after its chief rival, OpenAI, disclosed a similar incident.

Anthropic said in a blog post Thursday that it made the discovery after performing a review of its own cybersecurity tests, following OpenAI’s announcement of a breach. In both the OpenAI and Anthropic tests, the AI models were able to access the internet from within testing environments that should have been sealed off, according to Anthropic’s blog.

The company said it reviewed 141,006 evaluation tests and found three instances in which its Claude AI tool accessed the internet and then hacked into “the real-world infrastructure of external organizations.” The earliest incidents date to April, the company said.

The affected organizations aren’t named in the blog.

When Anthropic’s models gained unauthorized access to the three organizations, it treated each of them as part of an exercise. The tests were “capture-the-flag” evaluations in which the models sought hidden information by breaching other systems, a common way to test hacking capabilities for both humans and AI. However, the older model continued its attack even after getting evidence that it was running on the open internet; Anthropic’s most recent model stopped after recognizing it was on the internet, according to the blog.

The spate of accidental AI-caused hacks is already prompting some politicians to call for federal guardrails or other oversight of AI technology. More than 1,100 staffers across artificial intelligence firms also signed a petition on Tuesday, as Bloomberg first reported, that calls on the US government to support a mechanism that would help “deliberately pace” AI development to prevent the technology from advancing too fast.

Neither Anthropic nor the organizations that were breached had noticed the intrusions. In its blog, Anthropic said it could have done more to review network logs and evaluation transcripts.

The company disclosed the breaches nearly four months after it announced that it had developed a new AI model known as Mythos that was so powerful and potentially dangerous that the company strictly limited its release.

The breaches involved three different Claude models: Opus 4.7, Mythos 5 and an internal research test model, according to the blog. Each of the models ran without the safeguards normally implemented in public tools. Claude compromised the organizations using basic techniques such as exploiting weak passwords, according to the blog.

The incidents all occurred when Anthropic was using evaluation environments built by the AI security firm Irregular. In every case, Anthropic specified to Claude that its environment was a simulation and that it had no internet access.

“Due to a misunderstanding between us and our evaluation partner, this was not the case,” the blog says. An Irregular spokesperson said the company appreciates Anthropic’s collaboration and transparency. The company’s investigation is ongoing, the spokesperson said.

Anthropic said it regularly runs tests that involve simulating real cybersecurity challenges, calling them critical steps in developing and releasing models. However, the company also said it drew several lessons from the incidents, including that tests involving powerful autonomous capabilities also require significant controls.

“Safety testing happens before a model is released precisely because we don’t yet know what it is capable of,” the company said in its blog. “Evaluation environments increasingly need to be held to the same security standard as any other system our models run in.”

Anthropic was honest about the human failings that contributed to the hack, said Alan Woodward, professor of cybersecurity at the University of Surrey. The AI “hasn’t gone rogue — you’ve asked it to do something and left the gate open,” he said. “I think Anthropic were admitting that.”

— By Patrick Howell O’Neill and Shirin Ghaffary (Bloomberg News)

Tags: Anthropicartificial intelligence (AI)BloombergcybersecurityNewsPremium
Previous Post

Robinhood reports $100M in AUC for agentic AI trading tool

Next Post

Mastercard most concerned about frontier model cybersecurity threats

Related Posts

Fortinet headquarters
Risk & Security

Fortinet’s billing for AI-driven security operations grows 25% in Q2

July 30, 2026
humans and AI work together to pinpoint risk and suspicious activity
Risk & Security

Retaining the human component as AI combats fraud

July 28, 2026
a digital grid superimposed over a globe
Risk & Security

AI finding twice as many cyber flaws in 2026 as it did in 2025

July 27, 2026
Next Post
Signage for Mastercard during the Singapore FinTech Festival in Singapore, on Thursday, Nov. 3, 2022. The conference runs through Nov. 4. Photographer: Lionel Ng/Bloomberg

Mastercard most concerned about frontier model cybersecurity threats

EMERGING FINTECH DIRECTORY

Emerging Fintech Directory

The Buzz Podcast

SPONSORED

Build an Antifragile Strategy to Outperform the Market

July 14, 2026

How AI and Product Experts Turn Fuzzy Requirements Into Focused Dev-ready Roadmaps

April 19, 2026

Is Your Technology Supplier There for You?

April 1, 2026

  • About Us
  • Help Center
  • Contact Us
  • Privacy Terms
  • ADA Compliance
  • Advertise

 [wt_cli_manage_consent]

Connect

twitter linkedin podcast podcast podcast
© 2026 Royal Media
No Result
View All Result
  • NEWS
    • All News
    • Banking
    • Lending
    • Payments
    • Risk & Security
    • Strategy
  • AI News Tool [Beta]
  • DATA
  • TRANSACTIONS
  • EVENTS
    • FinAi Banking Summit
    • FinAi Lending Summit
  • PODCAST
  • WEBINARS
    • Webinar Library
  • SUBSCRIBE
  • Log In / Account

Welcome Back!

Login to your account below

Forgotten Password?

Retrieve your password

Please enter your username or email address to reset your password.

Log In

Unlock This Article

Create your free FinAi News account to access this article and stay informed on how AI is transforming financial services including banking, lending, payments, and risk.

Yes, I'd like to receive FinAi News updates, breaking news, and exclusive AI insights for financial services leaders.

Continue Reading with FinAi News Premium - Less than $2/Day

Upgrade to FinAi News Premium for unlimited access to news, insights, trends, and intelligence on how AI is transforming financial services including banking, lending, payments, and risk.
Upgrade to FinAi News Premium Subscription
No Result
View All Result
  • NEWS
    • All News
    • Banking
    • Lending
    • Payments
    • Risk & Security
    • Strategy
  • AI News Tool [Beta]
  • DATA
  • TRANSACTIONS
  • EVENTS
    • FinAi Banking Summit
    • FinAi Lending Summit
  • PODCAST
  • WEBINARS
    • Webinar Library
  • SUBSCRIBE
  • Log In / Account