Anthropic says its AI models hacked 3 organizations during testing

Advertisement

Advertise with us

Anthropic said its artificial intelligence models hacked into three other organizations during testing, just days after ChatGPT maker OpenAI raised concerns over AI controls after it disclosed its rogue models hacked another company.

Read this article for free:


or

Already have an account? Log in here »

To continue reading, please subscribe:

Subscribe and receive a limited-edition Free Press branded hat or tote.

Digital Subscription

One year of digital access for only $205*

  • Enjoy unlimited reading on winnipegfreepress.com
  • Read the E-Edition, our digital replica newspaper
  • Access News Break, our award-winning app
  • Play interactive puzzles

*First annual payment billed as $205.00 + GST for one year. This annual subscription will automatically renew at $233.00 + GST every 52 weeks (10% off the regular annual price of $259.35). Offer available to new and qualified returning subscribers only. Cancel any time.

To continue reading, please subscribe:

Add Free Press access to your Brandon Sun subscription for only an additional

$1 for the first 4 weeks*

  • Enjoy unlimited reading on winnipegfreepress.com
  • Read the E-Edition, our digital replica newspaper
  • Access News Break, our award-winning app
  • Play interactive puzzles
Start now

*Your next Brandon Sun subscription payment will increase by $1.00 and you will be charged $17.95 plus GST for four weeks. After four weeks, your payment will increase to $24.95 plus GST every four weeks.

Anthropic said its artificial intelligence models hacked into three other organizations during testing, just days after ChatGPT maker OpenAI raised concerns over AI controls after it disclosed its rogue models hacked another company.

Anthropic, the San Francisco-based AI company behind Claude, posted on its website Thursday that it discovered the three incidents after reviewing more than 141,000 evaluation runs.

It had launched a “large-scale” cybersecurity review which specifically looked for evidence whether its AI models were able to access the internet from within testing environments that should have been sealed off, in response to the OpenAI incident, Anthropic said.

FILE - Pages from the Anthropic website and the company's logo are displayed on a computer screen in New York, Feb. 26, 2026. (AP Photo/Patrick Sison, File)
FILE - Pages from the Anthropic website and the company's logo are displayed on a computer screen in New York, Feb. 26, 2026. (AP Photo/Patrick Sison, File)

Anthropic said the models involved in the incidents were Claude Opus 4.7, Claude Mythos 5 and an internal research test model. The earliest incidents date to April, the AI company said.

“Claude compromised the impacted organizations’ infrastructure using basic techniques,” Anthropic said, such as exploiting weak passwords.

In all three incidents, the AI models were tasked with a “capture the flag” cybersecurity challenge, which Anthropic said has been one of the ways it assesses a model’s cyber capabilities.

The models were given a fictional scenario and told a piece of secret information, or the “flag,” had been hidden on a different machine on the network with the objective of breaking in and retrieving it, it said.

It added that it had already reached out to the affected organizations, which it did not name. Two of them said they had not previously detected the activity. Anthropic said it was “continuing to reach out to the third.”

Anthropic said it conducted its review with Irregular, which describes itself as the “first frontier security lab.”

“Addressing these risks will require closer cooperation across the AI ecosystem,” Irregular said in a post on X.

Last week, OpenAI said its AI models went rogue during an evaluation of its models, breaking into the servers of AI startup Hugging Face. OpenAI described it as a “significant security incident.”

These incidents have highlighted the vulnerabilities in AI security and controls and raised questions over how AI can be safely kept under human control as the technology’s usage becomes more widespread globally.

Researchers have warned for years about risks from technology and the need for stronger AI defensive engineering.

“Safety testing happens before a model is released precisely because we don’t yet know what it is capable of,” Anthropic said on Thursday on its website.

Kok Tin Gan, co-founder & CEO of cybersecurity firm NyxLab, which specializes in cybersecurity and threat detection, believes there will be more such incidents in the future.

“It is increasingly about governing what agents are available to the AI, what authorities they possess, which actions require approval, and how we ensure they remain within scope,” Gan said.

But the future of AI safety extends beyond just the safety of AI models, he said.

“If we simply give the AI a goal and allow it to decide how to achieve it, we should not be surprised when it takes actions that technically satisfy the objective, but fall outside our intended scope or expectations,” Gan said.

Therefore, stepping up the governance of the organizations and authorities behind these AI models is going to be increasingly important, he said.

Report Error Submit a Tip

More Stories

Province takes aim at ER wait times by adding community-care capacity to free up hospital beds

Chris Kitching 5 minute read Preview

Province takes aim at ER wait times by adding community-care capacity to free up hospital beds

Chris Kitching 5 minute read Tuesday, Oct. 6, 2026

The province is aiming to free up hospital beds and cut wait times by adding 103 community-based care spaces in Winnipeg for patients who need shelter or specialized supports in order to be discharged.

Read
Tuesday, Oct. 6, 2026

Klein claims he was offered PC party position to exit 2022 mayoral race to help Gillingham win

Carol Sanders 6 minute read Preview

Klein claims he was offered PC party position to exit 2022 mayoral race to help Gillingham win

Carol Sanders 6 minute read Updated: Yesterday at 12:32 PM CDT

Mayoral candidate Kevin Klein said he was offered a position with the provincial Tories if he pulled out of the 2022 mayoral race that saw then-city councillor Scott Gillingham defeat former mayor Glen Murray.

Read
Updated: Yesterday at 12:32 PM CDT

Letters, Oct. 8

6 minute read Preview

Letters, Oct. 8

6 minute read 2:01 AM CDT

Computer scientist Melvin Conway once stated, “Any organization that designs a system (defined broadly) will produce a design whose structure is a copy of the organization’s communication structure.”

Read
2:01 AM CDT

Booze sales down, gaming and cannabis up in Manitoba

Carol Sanders 5 minute read Preview

Booze sales down, gaming and cannabis up in Manitoba

Carol Sanders 5 minute read 2:01 AM CDT

MANITOBA Liquor and Lotteries delivered $724.1 million in profits to the province in 2025-26, with the lion’s share coming from alcohol, even as booze sales continue to shrink.

The Crown corporation’s contribution to Manitoba’s general revenues decreased by $6 million or 0.8 per cent from last year’s allocation of $730.1 million, its annual report shows.

That indicates stable overall performance “despite a changing operating environment,” the recently-released report said.

Those changes include the order from Premier Wab Kinew last year to remove U.S. alcohol from Liquor Mart store shelves after President Donald Trump launched a trade war against Canada as well as Manitobans’ changing tastes.

Read
2:01 AM CDT

Duo accused in random Osborne Village attacks no strangers to police

Dean Pritchard and Chris Kitching 5 minute read Preview

Duo accused in random Osborne Village attacks no strangers to police

Dean Pritchard and Chris Kitching 5 minute read Tuesday, Oct. 6, 2026

A man arrested following a violent, random attack on two men in Osborne Village late Sunday night was released from jail a week earlier after a dozen charges, including crimes of violence, were stayed against him, court records show.

A man in his 40s was sent to hospital in critical condition and another man in his 20s in stable condition following separate attacks shortly before midnight in the vicinity of River Avenue and Osborne Street.

Police have charged Reginald Owen, 26, and Meaghan Owen, 32, both from Little Grand Rapids, with assault with a weapon and aggravated assault. They both remain in custody.

Court records show Reginald Owen was sentenced Sept. 24 to 30 days time served after pleading guilty to simple possession of magic mushrooms. Owen remained in custody until the following day when he was set to apply for bail on charges including assault, sexual assault, choking to overcome resistance, pointing a firearm, and carrying a concealed weapon.

Read
Tuesday, Oct. 6, 2026

Ladies and gentlemen of voting age, please turn your attention to the centre ring…

Dan Lett 5 minute read Preview

Ladies and gentlemen of voting age, please turn your attention to the centre ring…

Dan Lett 5 minute read Yesterday at 3:40 PM CDT

If there’s one thing that we in the news business wish for any election, it’s that the campaign is competitive and voters have multiple, viable options when it comes time to cast their ballots.

That is certainly what I wished for in this year’s Winnipeg mayoral election: a ballot with two or more competent candidates — including incumbent Scott Gillingham — that offered contrasting visions for the future of the city.

Unfortunately, this time we didn’t get competent or viable. Instead, we got a clown show stuffed with candidates for mayor who have demonstrated a stunning capacity for self-inflicted wounds.

It all started when mayoral candidates Kevin Klein and Mike Vogiatzakis both alleged they had been asked by each other’s campaign to step down to give the remaining candidate a better shot at defeating Gillingham.

Read
Yesterday at 3:40 PM CDT