Anthropic says its AI models hacked 3 organizations during testing

Advertisement

Advertise with us

Anthropic said its artificial intelligence models hacked into three other organizations during testing, just days after ChatGPT maker OpenAI raised concerns over AI controls after it disclosed its rogue models hacked another company.

Read this article for free:


or

Already have an account? Log in here »

To continue reading, please subscribe:

Subscribe and receive a limited-edition Free Press branded hat or tote.

Digital Subscription

One year of digital access for only $205*

  • Enjoy unlimited reading on winnipegfreepress.com
  • Read the E-Edition, our digital replica newspaper
  • Access News Break, our award-winning app
  • Play interactive puzzles

*First annual payment billed as $205.00 + GST for one year. This annual subscription will automatically renew at $233.00 + GST every 52 weeks (10% off the regular annual price of $259.35). Offer available to new and qualified returning subscribers only. Cancel any time.

To continue reading, please subscribe:

Add Free Press access to your Brandon Sun subscription for only an additional

$1 for the first 4 weeks*

  • Enjoy unlimited reading on winnipegfreepress.com
  • Read the E-Edition, our digital replica newspaper
  • Access News Break, our award-winning app
  • Play interactive puzzles
Start now

*Your next Brandon Sun subscription payment will increase by $1.00 and you will be charged $17.95 plus GST for four weeks. After four weeks, your payment will increase to $24.95 plus GST every four weeks.

Anthropic said its artificial intelligence models hacked into three other organizations during testing, just days after ChatGPT maker OpenAI raised concerns over AI controls after it disclosed its rogue models hacked another company.

Anthropic, the San Francisco-based AI company behind Claude, posted on its website Thursday that it discovered the three incidents after reviewing more than 141,000 evaluation runs.

It had launched a “large-scale” cybersecurity review which specifically looked for evidence whether its AI models were able to access the internet from within testing environments that should have been sealed off, in response to the OpenAI incident, Anthropic said.

FILE - Pages from the Anthropic website and the company's logo are displayed on a computer screen in New York, Feb. 26, 2026. (AP Photo/Patrick Sison, File)
FILE - Pages from the Anthropic website and the company's logo are displayed on a computer screen in New York, Feb. 26, 2026. (AP Photo/Patrick Sison, File)

Anthropic said the models involved in the incidents were Claude Opus 4.7, Claude Mythos 5 and an internal research test model. The earliest incidents date to April, the AI company said.

“Claude compromised the impacted organizations’ infrastructure using basic techniques,” Anthropic said, such as exploiting weak passwords.

In all three incidents, the AI models were tasked with a “capture the flag” cybersecurity challenge, which Anthropic said has been one of the ways it assesses a model’s cyber capabilities.

The models were given a fictional scenario and told a piece of secret information, or the “flag,” had been hidden on a different machine on the network with the objective of breaking in and retrieving it, it said.

It added that it had already reached out to the affected organizations, which it did not name. Two of them said they had not previously detected the activity. Anthropic said it was “continuing to reach out to the third.”

Anthropic said it conducted its review with Irregular, which describes itself as the “first frontier security lab.”

“Addressing these risks will require closer cooperation across the AI ecosystem,” Irregular said in a post on X.

Last week, OpenAI said its AI models went rogue during an evaluation of its models, breaking into the servers of AI startup Hugging Face. OpenAI described it as a “significant security incident.”

These incidents have highlighted the vulnerabilities in AI security and controls and raised questions over how AI can be safely kept under human control as the technology’s usage becomes more widespread globally.

Researchers have warned for years about risks from technology and the need for stronger AI defensive engineering.

“Safety testing happens before a model is released precisely because we don’t yet know what it is capable of,” Anthropic said on Thursday on its website.

Kok Tin Gan, co-founder & CEO of cybersecurity firm NyxLab, which specializes in cybersecurity and threat detection, believes there will be more such incidents in the future.

“It is increasingly about governing what agents are available to the AI, what authorities they possess, which actions require approval, and how we ensure they remain within scope,” Gan said.

But the future of AI safety extends beyond just the safety of AI models, he said.

“If we simply give the AI a goal and allow it to decide how to achieve it, we should not be surprised when it takes actions that technically satisfy the objective, but fall outside our intended scope or expectations,” Gan said.

Therefore, stepping up the governance of the organizations and authorities behind these AI models is going to be increasingly important, he said.

Report Error Submit a Tip

More Stories

Manitoba abandoning spring, fall clock changes

Carol Sanders 8 minute read Preview

Manitoba abandoning spring, fall clock changes

Carol Sanders 8 minute read Updated: Yesterday at 5:36 PM CDT

Time’s up for the time change in Manitoba.

Manitobans won’t be setting their clocks back an hour in November. The province has decided to permanently stay on daylight time after surveying residents on the twice annual clock adjustment.

“We heard from Manitobans loud and clear — no more changing the clocks in Manitoba,” Premier Wab Kinew said in a news release Thursday.

Most of the 72,000 votes submitted in an online EngageMB survey — 92 per cent — said they would like to see the practice end. The survey asked if they preferred earlier sunrises in winter with standard time or the later sunsets in summer with daylight time. More than half of survey participants (58 per cent) indicated a preference to stay on permanent daylight time; 34 per cent voted to eliminate the time change but stay on standard time.

Read
Updated: Yesterday at 5:36 PM CDT

What to watch as puck drops on Jets training camp

Mike McIntyre and Ken Wiebe 10 minute read Preview

What to watch as puck drops on Jets training camp

Mike McIntyre and Ken Wiebe 10 minute read Wednesday, Sep. 16, 2026

Welcome to a Winnipeg Jets training camp like no other — one where the on-ice storylines and battles threaten to be trumped by what’s happening away from the rink.

Yes, the dark cloud that is the Connor Hellebuyck situation looms over the organization as it prepares to try and turn the page on a terrible 2025-26 campaign which ended with a 26th-overall finish.

It’s not going to be easy, especially in a highly competitive Central Division and Western Conference and with such a major unresolved issue.

But the work starts now, and Free Press hockey writers Mike McIntyre and Ken Wiebe get you set for what to expect ahead of the 84-game regular season, which begins Oct. 2.

Read
Wednesday, Sep. 16, 2026

Police officer faces charges of domestic assault

Free Press staff 2 minute read Preview

Police officer faces charges of domestic assault

Free Press staff 2 minute read Yesterday at 4:26 PM CDT

A Winnipeg police officer has been charged after police said he broke into a woman’s home and attacked her.

Police were sent to a home on Wednesday shortly before 10:20 p.m., the Winnipeg Police Service said in a Thursday news release. When officers arrived, they found the man, 20-year member in his 40s, inside a home assaulting the woman, the release said.

The man was taken into custody and the woman was given medical aid.

She told officers the man broke into the home, threatened to kill her and then attacked her.

Read
Yesterday at 4:26 PM CDT

Jets GM insists club’s not working under any deadline in netminder’s trade request

Mike McIntyre 7 minute read Preview

Jets GM insists club’s not working under any deadline in netminder’s trade request

Mike McIntyre 7 minute read Updated: Yesterday at 5:27 PM CDT

Connor Hellebuyck and his family were subject to so much post-Olympic vitriol that both the Winnipeg Jets security team and even police had to get involved.

Read
Updated: Yesterday at 5:27 PM CDT

Jets vets weigh in on franchise goalie Hellebuyck not reporting to training camp

Mike McIntyre 7 minute read Preview

Jets vets weigh in on franchise goalie Hellebuyck not reporting to training camp

Mike McIntyre 7 minute read Yesterday at 6:35 PM CDT

Hockey is the ultimate team sport, where it’s supposed to be about the crest on your chest rather than the name on your back.

So how is Connor Hellebuyck’s trade request and subsequent training camp strike playing out inside the Winnipeg Jets locker room — particularly among other core players he’s gone to battle with over the years?

“Obviously the news is heartbreaking,” Jets forward Kyle Connor said Thursday.

“It feels like we had such a special group here and for whatever reason — I can’t speak to the reason until we hear from the man, I’m not going to speculate on what it is but I know that a lot of guys in that room that are heartbroken.”

Read
Yesterday at 6:35 PM CDT

‘She was looking to destroy me’: convicted harasser faces new charges

Dean Pritchard 6 minute read Preview

‘She was looking to destroy me’: convicted harasser faces new charges

Dean Pritchard 6 minute read Yesterday at 6:54 PM CDT

When Agnieszka Ciochon-Newton moved into the 55-plus apartment building on Stradbrook Avenue last year, she came across as “timid, soft-spoken and sensitive,” said building manager Tanya Owen.

Little did Owen know the meek-looking 58-year-old woman would soon turn her life and the lives of many of her tenants upside down and threaten her career and family.

“She was looking to destroy me for a simple tenant disagreement,” Owen said Thursday.

Court records show Ciochon-Newton was arrested earlier this month and charged with seven counts of criminal harassment, two counts of uttering threats and one count each of fraud, breaking and entering, attempting to obstruct justice, intimidation of a justice official, mischief and defamatory libel.

Read
Yesterday at 6:54 PM CDT