FinAi News

No products in the cart.

Subscribe
  • News
  • AI News Tool
  • Data
  • Transactions
  • Events
    • FinAi Banking Summit
    • FinAi Lending Summit
  • Podcast
  • WEBINARS
    • Webinar Library
Log In
No Result
View All Result
  • Banking
  • Lending
  • Payments
  • Risk & Security
  • Strategy
FinAi News
  • News
  • AI News Tool
  • Data
  • Transactions
  • Events
    • FinAi Banking Summit
    • FinAi Lending Summit
  • Podcast
  • WEBINARS
    • Webinar Library
BAN PLUS
Log In
No Result
View All Result
FinAi News
No Result
View All Result

OpenAI, Anthropic model tests reveal more hacking

Mythos carried out 17 of 19 identified 'autonomous, unsanctioned actions taken on the internet'

Bloomberg NewsbyBloomberg News
August 5, 2026
in Risk & Security
Reading Time: 2 mins read
0
Share on Facebook

AI models developed by OpenAI and Anthropic PBC carried out “unsanctioned” actions — including hacking a website and attempting to inject harmful code into software during safety testing — reinforcing fears that neither the creators nor seasoned researchers of these systems can predict their actions in testing.

The U.K. government’s AI Security Institute, established in 2023 to evaluate the safety of cutting-edge AI models, said Tuesday that both Anthropic’s Mythos 5 and OpenAI’s GPT-5.6-Sol models had “engaged in sustained, potentially harmful activity directed at real people and organizations” during evaluations. The institute intentionally allowed the models internet access and used them without certain safety filters to test their capabilities.

“Even under test conditions, this incident is significant: It is the first time we have seen risks around autonomy and deception manifest this clearly in the real world,” the institute said in a post on the social media platform X.

In one instance, the testing organization said, Mythos 5 attempted to add harmful code to an open-source software project on GitHub. It went as far as to create fake identities in an effort to get its code approved. “A human maintainer caught and refused to approve the malicious code,” the group wrote.

Over the past two weeks, both OpenAI and Anthropic have publicly acknowledged that they’ve collectively breached the systems of multiple institutions, including Hugging Face Inc., inadvertently while testing their models. The latest disclosures serve as fresh evidence that AI agents are capable of acting autonomously in ways that even researchers trained to root out vulnerabilities in the technology can no longer anticipate, underscoring the need for both more rigorous safety screening and more foolproof testing environments.

The U.K.’s AI Security Institute said Anthropic’s Mythos 5 model carried out 17 of the 19 “autonomous, unsanctioned actions taken on the internet“ that it detected.

Some U.S. government leaders have called for more oversight of the technology in response to the breaches. Last Tuesday, more than 1,100 AI industry workers signed a petition pushing for a regulatory mechanism that would “deliberately pace” AI technology and prevent it from advancing too quickly.

Anthropic said on X that it’s working with the U.K. security institute to “gather more details of the incident as we conduct our own investigation.”

ChatGPT maker OpenAI separately flagged in a blog post that yet another security incident occurred during the testing of one of its models with Irregular, an external cybersecurity firm.

In this incident, OpenAI’s models were subject to a so-called capture-the-flag test in which they were tasked with finding information hidden in a simulated environment. The models took advantage of a “misconfiguration” in the testing environment to connect to the internet and hack the website of an unidentified institution, the company said.

The breach occurred when OpenAI models were undergoing the same Irregular evaluation that resulted in Anthropic’s models hacking three organizations, a person familiar with the matter said, asking not to be identified because the information isn’t public. Anthropic disclosed those breaches last week.

An Irregular spokesperson declined to comment.

Two weeks ago, OpenAI disclosed that its models were behind an unprecedented hack against the startup Hugging Face. In that incident, the models exploited a vulnerability to “escape” their sandbox testing environment and connect to the internet, at which point they breached Hugging Face’s system, which hosts AI models and datasets.

Tags: Anthropicartificial intelligence (AI)BloomberghackingNewsOpenAIPremium
Previous Post

Banks rethink vendor outsourcing amid rapid AI advancement

Related Posts

UBS corporate headquarters
Risk & Security

FinCEN assesses $125M penalty against UBS

August 4, 2026
man decides whether or not to accept call identified as fraud
Risk & Security

FIs need to educate consumers about AI-supported fraud

August 3, 2026
(Courtesy/Bloomberg)
Risk & Security

Anthropic AI models hacked three organizations during tests

July 31, 2026

EMERGING FINTECH DIRECTORY

Emerging Fintech Directory

FinAi Podcast

SPONSORED

Build an Antifragile Strategy to Outperform the Market

July 14, 2026

How AI and Product Experts Turn Fuzzy Requirements Into Focused Dev-ready Roadmaps

April 19, 2026

Is Your Technology Supplier There for You?

April 1, 2026

  • About Us
  • Help Center
  • Contact Us
  • Privacy Terms
  • ADA Compliance
  • Advertise

 [wt_cli_manage_consent]

Connect

twitter linkedin podcast podcast podcast
© 2026 Royal Media
No Result
View All Result
  • NEWS
    • All News
    • Banking
    • Lending
    • Payments
    • Risk & Security
    • Strategy
  • AI News Tool [Beta]
  • DATA
  • TRANSACTIONS
  • EVENTS
    • FinAi Banking Summit
    • FinAi Lending Summit
  • PODCAST
  • WEBINARS
    • Webinar Library
  • SUBSCRIBE
  • Log In / Account

Welcome Back!

Login to your account below

Forgotten Password?

Retrieve your password

Please enter your username or email address to reset your password.

Log In

Unlock This Article

Create your free FinAi News account to access this article and stay informed on how AI is transforming financial services including banking, lending, payments, and risk.

Yes, I'd like to receive FinAi News updates, breaking news, and exclusive AI insights for financial services leaders.

Continue Reading with FinAi News Premium - Less than $2/Day

Upgrade to FinAi News Premium for unlimited access to news, insights, trends, and intelligence on how AI is transforming financial services including banking, lending, payments, and risk.
Upgrade to FinAi News Premium Subscription
No Result
View All Result
  • NEWS
    • All News
    • Banking
    • Lending
    • Payments
    • Risk & Security
    • Strategy
  • AI News Tool [Beta]
  • DATA
  • TRANSACTIONS
  • EVENTS
    • FinAi Banking Summit
    • FinAi Lending Summit
  • PODCAST
  • WEBINARS
    • Webinar Library
  • SUBSCRIBE
  • Log In / Account