Will AI models achieve the ability to improve autonomously? Leading labs say the scenario is near

Advertisement

Advertise with us

Once a distant ambition for technology researchers, the prospect of artificial intelligence models teaching themselves autonomously to be more efficient and capable appears ever closer to reality.

Read this article for free:


or

Already have an account? Log in here »

To continue reading, please subscribe:

Subscribe and receive a limited-edition Free Press branded hat or tote.

Digital Subscription

One year of digital access for only $205*

  • Enjoy unlimited reading on winnipegfreepress.com
  • Read the E-Edition, our digital replica newspaper
  • Access News Break, our award-winning app
  • Play interactive puzzles

*First annual payment billed as $205.00 + GST for one year. This annual subscription will automatically renew at $233.00 + GST every 52 weeks (10% off the regular annual price of $259.35). Offer available to new and qualified returning subscribers only. Cancel any time.

To continue reading, please subscribe:

Add Free Press access to your Brandon Sun subscription for only an additional

$1 for the first 4 weeks*

  • Enjoy unlimited reading on winnipegfreepress.com
  • Read the E-Edition, our digital replica newspaper
  • Access News Break, our award-winning app
  • Play interactive puzzles
Start now

*Your next Brandon Sun subscription payment will increase by $1.00 and you will be charged $17.95 plus GST for four weeks. After four weeks, your payment will increase to $24.95 plus GST every four weeks.

Once a distant ambition for technology researchers, the prospect of artificial intelligence models teaching themselves autonomously to be more efficient and capable appears ever closer to reality.

As the technology advances, developers say it is approaching “recursive self-improvement,” or RSI, in which AI models find ways to improve themselves and build their successor. It could bring the promise of advances in science and medicine, tech company executives say, but also risks.

The uncertainty over where it all could lead is at the heart of growing fears about AI evading human control, and possible threats to humanity, which led several AI moguls to join last weekend in a call to slow down the technology’s pace of growth.

FILE - Dario Amodei, CEO and co-founder of Anthropic, attends the annual meeting of the World Economic Forum in Davos, Switzerland, Jan. 23, 2025. (AP Photo/Markus Schreiber, File)
FILE - Dario Amodei, CEO and co-founder of Anthropic, attends the annual meeting of the World Economic Forum in Davos, Switzerland, Jan. 23, 2025. (AP Photo/Markus Schreiber, File)

Anthropic this week detailed how its model Claude is helping the company to develop the next, more intelligent version of itself. Claude is now leading 26% of Anthropic’s model research and development, which the company said means it can complete most of a given task “end-to-end from a high-level prompt” while still being under human supervision. The models are not working completely autonomously — at least not yet.

Here are some key points about recursive self-improvement.

What is recursive self-improvement?

Leading AI companies have different definitions for recursive self-improvement. Some define it as when there is any feedback from AI on model improvement, while others define it as AI working toward that goal fully autonomously.

Autonomous recursive self-improvement essentially means AI that can improve itself designing the next version of the system, then the next version, and so on, said Anthony Aguirre, president and CEO of the nonprofit Future of Life Institute and a physics professor at the University of California, Santa Cruz.

“The really important thing here is that as AI is doing more of it, it gets faster, because AI operates just much, much more quickly than the humans do,” he said.

The fear around RSI is based largely on a runaway superintelligence emerging from that process, said John Thickstun, an assistant professor of computer science at Cornell University who studies methods that control the behavior of AI models. But he said a more grounded view suggests a kind of recursive self-improvement has been going on in AI development for a while now.

“We have already, for years, been using these models in supportive roles for creating the next version of these models. So people use the past generation of models to write code for the AI systems that then create the next generation,” he said.

For years, prominent AI researchers such as OpenAI co-founder Andrej Karpathy have experimented with trying to get AI models to train and improve new AI systems. Those efforts have brought minor improvements, but not big creative leaps, Thickstun said.

But AI companies today, Aguirre said, are much closer to pulling off those bigger leaps in improvement.

“You can see in these plots from Anthropic over time, more and more of research is being done by the AI and it’s becoming closer and closer to fully autonomous,” he said. “And the result of that success, ultimately is something that is, I think, extremely scary. I think this is probably the worst idea in the history of humanity to do this. And yes, they’re doing it.”

Some AI labs say RSI is not far off

Anthropic’s recent announcement provided the public — and other labs — with some insight into RSI progress, and it encouraged its competitors to share similar metrics. Still, the company has not expressly said how close it is to achieving fully autonomous model improvement.

ChatGPT maker OpenAI announced this month that it has developed an automated “research intern,” which it defines as a system that can carry out well-defined research tasks under human direction, including “tasks that would take a skilled researcher a few days.” The company has said it is moving forward with the goal of creating an automated AI “researcher” by March 2028.

The company said in that announcement that while RSI can help align models’ actions with human values and intentions, that doesn’t mean “rapid RSI is necessarily an outcome we should pursue.”

“Whether and how to proceed must depend on our ability to preserve human control and on informed democratic choices about the benefits and risks,” the company said in a blog post.

Elon Musk seems more eager to forge ahead. He said in March that for xAI’s Grok models, “humans are gradually getting less and less in the loop” on model improvement and that “every successive model is built by the one before it,” but clarified that the process was not yet fully automated. That target might be reached by the end of this year, he added, “but not later” than 2027.

Microsoft and some other leading AI companies seem to be taking a different approach.

Mustafa Suleyman, the CEO of Microsoft AI, has said the company is moving toward “humanist superintelligence,” or advanced AI capabilities that are in service of people and humanity at large. Suleyman said in a 2025 essay that this would not mean “an unbounded and unlimited entity with high degrees of autonomy,” but rather AI that is “carefully calibrated, contextualized, within limits.”

How development slowdown talks could impact RSI

A key challenge labs face — and have been facing essentially since the technology’s inception — is ensuring their safety measures advance alongside the models’ capabilities.

Divisions have emerged in the tech industry over calls for a coordinated AI slowdown for safety, and not every major player in the AI space has specifically commented on their path forward with RSI.

Anthropic, which has been a leading voice in the calls for pacing, has said it would slow or temporarily pause its development work — assuming its global competitors also did so, and in a “verifiable manner.”

OpenAI explicitly said this month it does not yet know how to “safely get all the way to aligned, full RSI,” adding that the company “cannot assume that progress in alignment and safety will keep pace.” More capable systems can become harder to monitor, it continued, but pursuing RSI is still a goal it says it values because an “automated AI researcher can also be an automated safety or alignment researcher.”

Report Error Submit a Tip

More Stories

Jets GM insists club’s not working under any deadline in netminder’s trade request

Mike McIntyre 7 minute read Preview

Jets GM insists club’s not working under any deadline in netminder’s trade request

Mike McIntyre 7 minute read Thursday, Sep. 17, 2026

Connor Hellebuyck and his family were subject to so much post-Olympic vitriol that both the Winnipeg Jets security team and even police had to get involved.

Read
Thursday, Sep. 17, 2026

His blissful kiss is well worth the risk

Maureen Scurfield 4 minute read Yesterday at 2:00 AM CDT

DEAR MISS LONELYHEARTS: The best kisser I ever met knocked me off balance. My little joke was, “That’s why we have to go home now and lie down in a hurry!” We’ve been seeing each other for almost six months. This man can convince me to do almost anything — so much fun!

But, now he’s been transferred out of province for bigger money at a new company.

He’s leaving soon and he could have asked me to move with him, as his trucking partner, but he didn’t! Other women will soon have their hands on him, believe me!

I’m already 27 and I don’t want to date any more guys — just him. Now what can I do? I feel sadder every day.

Near-future U.S. the setting for surreal dual-world storyline

Reviewed by Morley Walker 4 minute read Preview

Near-future U.S. the setting for surreal dual-world storyline

Reviewed by Morley Walker 4 minute read 2:01 AM CDT

It’s not every day you encounter a seriously intended literary novel that attempts to dramatize a mind-bending physics theory like quantum entanglement.

But then the B.C.-born-and-raised Emily St. John Mandel is not your everyday Canadian literary writer.

In her late 40s, she has lived in the U.S. for her entire writing career. She is best known for her 2014 dystopian novel Station Eleven, about a Shakespearean acting troupe navigating a pandemic-ravaged America.

Her new outing, her seventh, Exit Party, shares DNA with detective and sci-fi genres more than with conventional literary fiction.

Read
2:01 AM CDT

Canada must not accept another Iranian autocracy

David Matas 5 minute read 2:01 AM CDT

Iran is a state run by a radical government victimizing its citizens and its neighbours. How do we change that? Any regime change that repeats the mistakes of the past is not a solution.

What went wrong in the past which led to the current regime? One answer is the U.K. and American aided coup which led to the ouster in 1953 of a democratically elected government headed at the time by Mohammad Mossadegh and its replacement with an American friendly autocrat, the Shah of Iran, Mohammad Reza Pahlavi.

The Shah imposed his rule through the Iranian National Intelligence and Security Organization — SAVAK, the acronym of its Persian name. SAVAK systematically inflicted torture, arbitrary killings and extra-judicial killings on perceived opponents of the regime. It was only a matter of time before repression of a regime directed to serving foreign interests would be overthrown.

What replaced the regime of the Shah, in 1979, the current regime of the mullahs, is a regime as intolerant and violent as the regime of the Shah and then some. That was not inevitable. But neither was it surprising. The vicious nature of the current regime, following the example set by the Shah, and its hatred for the United States, reacting to the U.S. aided imposition of the Shah, both echoes and reacts to the past. In 1979, the U.S. reaped what it sowed in 1953.

Not good bet to shoulder gambling issue alone

Maureen Scurfield 5 minute read 2:01 AM CDT

DEAR MISS LONELYHEARTS: We’re living a lie over at our house, but it’s holding the family together.

I know my husband is deeply into gambling. If it weren’t for my steady job I don’t know where the family would be! I insist on paying all of our bills. I also funnel some major cash off to my sister, in case my husband tries to tap into my accounts somehow when I’m out of the house.

He doesn’t know my passwords, since I keep changing them, but he sometimes shows up behind me when I’m online, like he’s dying to ask something.

I say, “What are you looking at? Stop looking over my shoulder!” Then he makes a lame joke and goes away.

Kinew delivered what most of us wanted; now we’ll have to deal with the consequences

Dan Lett 5 minute read Preview

Kinew delivered what most of us wanted; now we’ll have to deal with the consequences

Dan Lett 5 minute read Updated: Yesterday at 8:43 AM CDT

Premier Wab Kinew has decided to adopt permanent daylight savings time, which means a permanent end to the spring forward, fall back tradition of adjusting our clocks. And absolutely no one is surprised.

Read
Updated: Yesterday at 8:43 AM CDT