AI models’ breakout from human control brings a told-you-so moment for technology researchers

Advertisement

Advertise with us

It is the kind of development once seen only in science fiction: An artificial intelligence system, trained to probe for digital vulnerabilities, breaks free of human control and acts on its own to hack another company.

Read this article for free:


or

Already have an account? Log in here »

To continue reading, please subscribe:

Subscribe and receive a limited-edition Free Press branded hat or tote.

Digital Subscription

One year of digital access for only $205*

  • Enjoy unlimited reading on winnipegfreepress.com
  • Read the E-Edition, our digital replica newspaper
  • Access News Break, our award-winning app
  • Play interactive puzzles

*First annual payment billed as $205.00 + GST for one year. This annual subscription will automatically renew at $233.00 + GST every 52 weeks (10% off the regular annual price of $259.35). Offer available to new and qualified returning subscribers only. Cancel any time.

To continue reading, please subscribe:

Add Free Press access to your Brandon Sun subscription for only an additional

$1 for the first 4 weeks*

  • Enjoy unlimited reading on winnipegfreepress.com
  • Read the E-Edition, our digital replica newspaper
  • Access News Break, our award-winning app
  • Play interactive puzzles
Start now

*Your next Brandon Sun subscription payment will increase by $1.00 and you will be charged $17.95 plus GST for four weeks. After four weeks, your payment will increase to $24.95 plus GST every four weeks.

It is the kind of development once seen only in science fiction: An artificial intelligence system, trained to probe for digital vulnerabilities, breaks free of human control and acts on its own to hack another company.

The attack announced this week by OpenAI, which blamed rogue AI models, underscored the blistering growth in the technology’s capabilities. For many, it also added urgency to questions about whether and how it can be prevented from causing mayhem on a bigger scale, with more serious consequences.

In what OpenAI called an “unprecedented” episode, the company said its advanced AI models used stolen credentials to break into the servers of an AI startup. It started in what was supposed to be a “highly isolated” testing environment, with reduced guardrails, before the AI agent found its way onto the internet.

But the disclosure brought a told-you-so moment for researchers who have called for a slowdown of AI development and warned for years that the technology could pose existential risks to humanity. In its wake, experts have called for improved testing by the AI companies and more dialogue between the U.S. and China to come up with shared solutions.

“I think we’ve got to take this as a warning shot to not make them smarter, and that probably is going to require global collaboration,” said Nate Soares, co-author of the 2025 book “If Anyone Builds It, Everyone Dies.”

The hack could pressure companies to improve containment

If a model can decide to do something unethical, illegal or harmful on its own, what — if anything — can humans do to prevent it from doing so?

OpenAI said it had tasked the AI models involved with pursuing “advanced exploitation using complex attack paths” to test cyber capabilities, but the technology went to unexpected lengths. It apparently decided on its own to target Hugging Face, a well-known AI development hub and marketplace, to obtain information it needed to carry out a task.

Zahra Timsah, the co-founder and CEO of governance platform i-GENTIC AI, said she expects the incident to increase pressure on OpenAI and its competitors to complete rigorous testing and explore containment more thoroughly before AI systems are made accessible to the public.

Monitoring an agent’s behavior after the fact, as OpenAI is now doing with its investigation, is no longer enough, she said. “It’s like having a seat belt, air bags, brakes, everything in the car. It should be there before the car starts driving,” Timsah said.

The disclosure comes amid heightened concerns about the cybersecurity capabilities of powerful models. In June, President Donald Trump signed an executive order creating a framework for the federal government to vet the national security risks of the most advanced AI systems for up to a month before their public release.

OpenAI said Thursday that it briefed the White House this week about the Hugging Face attack.

Other experts see the event as a sign of AI’s growing pains

Some experts say the hack is part of the trial and error that comes with improving cybersecurity capabilities and is no cause for panic.

“We’ve been dealing with people creating cybersecurity attacks for as long as the internet has existed. And one of the interesting properties of these language models is that the same capabilities that make them able to perform cybersecurity attacks also allow them to do cybersecurity threat analysis and make cybersecurity defenses,” said John Thickstun, an assistant professor of computer science at Cornell University who studies methods that control the behavior of AI models.

The disclosure has raised skepticism from those who say it advantages OpenAI to make its technology seem scarier. Given that humans at OpenAI had decided to turn off some safeguards for the test, some have argued the outcome should not have been terribly surprising.

FILE - The OpenAI logo is displayed on a cell phone in front of an image generated by ChatGPT's Dall-E text-to-image model, Dec. 8, 2023, in Boston. (AP Photo/Michael Dwyer, File)
FILE - The OpenAI logo is displayed on a cell phone in front of an image generated by ChatGPT's Dall-E text-to-image model, Dec. 8, 2023, in Boston. (AP Photo/Michael Dwyer, File)

Thickstun noted the disclosure plays into the need of OpenAI, a startup working toward a Wall Street debut, to raise money.

“The story that they’ve been consistently telling over the lifetime of this company is a story about how dangerous their models are, which their investors read as a story of how powerful their language models are,” he said.

The disclosure renews calls for more regulation

The hack renewed calls in some corners for increased regulations and oversight of AI companies.

U.S. Rep. Greg Casar, a Texas Democrat, wrote on social media: “We need regular mandatory independent safety testing and oversight, mandatory disclosure of security incidents, and international cooperation to keep people safe from absolute disaster.”

Soares, director of the Machine Intelligence Research Institute, said the U.S. will need to open talks with its biggest AI competitor, China, something he thinks is not as outlandish as it might have seemed even a year ago. China’s leader Xi Jinping warned at a conference just last week of the need to keep AI from evading human control. And after an early aversion to regulating AI, Trump’s administration has grown more restrictive at reining in cybersecurity risks.

“A lot can change when the national security community starts to notice that they have a serious threat,” Soares said. “Will this wake them up? Hopefully. I’m not sure. If this doesn’t, maybe the next incident will.”

AI pioneer Yoshua Bengio said on social media the episode is deeply concerning and should serve as a “wake-up call.”

“Continuing on the current trajectory of AI development will likely lead to an increase in concrete cases of autonomous cyberattacks as well as other high-risk incidents of misaligned and dangerous AI behavior,” said Bengio, a professor at the University of Montreal. “We urgently need to take action to prevent these situations, rather than attempting to clean up the damage after the fact.”

Report Error Submit a Tip

More Stories

AI models’ breakout from human control brings a told-you-so moment for technology researchers

Barbara Ortutay, Kaitlyn Huamani And Matt O'brien, The Associated Press 6 minute read Preview

AI models’ breakout from human control brings a told-you-so moment for technology researchers

Barbara Ortutay, Kaitlyn Huamani And Matt O'brien, The Associated Press 6 minute read Monday, Aug. 10, 2026

It is the kind of development once seen only in science fiction: An artificial intelligence system, trained to probe for digital vulnerabilities, breaks free of human control and acts on its own to hack another company.

The attack announced this week by OpenAI, which blamed rogue AI models, underscored the blistering growth in the technology’s capabilities. For many, it also added urgency to questions about whether and how it can be prevented from causing mayhem on a bigger scale, with more serious consequences.

In what OpenAI called an “unprecedented” episode, the company said its advanced AI models used stolen credentials to break into the servers of an AI startup. It started in what was supposed to be a “highly isolated” testing environment, with reduced guardrails, before the AI agent found its way onto the internet.

But the disclosure brought a told-you-so moment for researchers who have called for a slowdown of AI development and warned for years that the technology could pose existential risks to humanity. In its wake, experts have called for improved testing by the AI companies and more dialogue between the U.S. and China to come up with shared solutions.

Read
Monday, Aug. 10, 2026

RCMP searching for missing woman from Split Lake

1 minute read Saturday, Aug. 15, 2026

Thompson RCMP are asking for help from the public to find a 37-year-old woman who has not been heard from since late July.

Melanie Beardy, from Split Lake First Nation, was last in contact on July 26, said an RCMP news release Saturday. Beardy, who was in Thompson, was reported missing last Sunday and RCMP said both police and her family are worried for her well-being.

Attempts to find Beardy have been unsuccessful, said RCMP.

She’s described as 5’4 with a large build and long, black hair usually worn in a ponytail, said Mounties. She has a tattoo on her chest that reads: “Elvis,” and has a tattoo of a dreamcatcher on her left forearm.

Puzzles Palace

1 minute read Monday, Jul. 27, 2026

To solve our puzzles, please subscribe with this special offer: |

Hometown hero hits CEBL championship clinching shot

Taylor Allen 4 minute read Preview

Hometown hero hits CEBL championship clinching shot

Taylor Allen 4 minute read Yesterday at 6:34 PM CDT

Chapter 1 of Winnipeg Sea Bears basketball was capped by hometown hero Chad Posthumus.

It was May 27, 2023, and the late, great centre out of River East Collegiate had a put-back basket off a Teddy Allen miss to give the Sea Bears a 90-85 victory over the Vancouver Bandits in Winnipeg’s inaugural game, in front of what was then a record crowd of 7,303.

So, on Saturday, it was rather poetic that Simon Hildebrandt — another Winnipegger and a player Posthumus used to mentor — was the one to put an exclamation mark on the franchise’s latest and greatest chapter.

Hildebrandt drilled a wide-open three in Target Score Time to crown the Sea Bears as the 2026 CEBL champions.

Read
Yesterday at 6:34 PM CDT

Manitoba family finalist for cattle producers environmental award

Aaron Epp 4 minute read Preview

Manitoba family finalist for cattle producers environmental award

Aaron Epp 4 minute read Saturday, Aug. 15, 2026

A family in western Manitoba is one of seven finalists for a national environmental award honouring beef cattle producers.

Connor and Kyla English, alongside Connor’s parents, Brian and Leanne English, are nominated for the Canadian Cattle Association’s Environmental Stewardship Award. The 30-year-old award is presented to beef farmers and ranchers who explore innovative ways to protect, preserve and enhance their operations and the environment simultaneously.

This year’s recipient will be announced on Wednesday during a banquet at the Canadian Beef Industry Conference in Winnipeg.

The English family’s roots near Bradwardine — about 265 kilometres west of Winnipeg — date back to 1898. Today, Connor, Kyla, Brian and Leanne manage a herd of about 320 breeding cows on nearly 445 hectares (1,100 acres) of owned and rented pasture and forage land.

Read
Saturday, Aug. 15, 2026

Today’s horoscope

Georgia Nicols 4 minute read Preview

Today’s horoscope

Georgia Nicols 4 minute read 2:00 AM CDT

MOON ALERT: Caution. Avoid shopping (except food and gas) or making important decisions from 6 a.m. until 5 p.m. After that, the moon moves from Libra into Scorpio.

ARIES (March 21-April 19)

This is a fun-loving, playful day. Enjoy shmoozing with everyone. Sports and fun times with kids will appeal. Be prepared to co-operate with others because you might have to go more than halfway when dealing with spouses, partners and close friends. Romantic vibes are strong.

TAURUS (April 20-May 20)

Read
2:00 AM CDT