How do you know when AI is powerful enough to be dangerous? Regulators try to do the math

Advertisement

Advertise with us

How do you know if an artificial intelligence system is so powerful that it poses a security danger and shouldn’t be unleashed without careful oversight?

Read this article for free:


or

Already have an account? Log in here »

To continue reading, please subscribe:

Subscribe and receive a limited-edition Free Press branded hat or tote.

Digital Subscription

One year of digital access for only $205*

  • Enjoy unlimited reading on winnipegfreepress.com
  • Read the E-Edition, our digital replica newspaper
  • Access News Break, our award-winning app
  • Play interactive puzzles

*First annual payment billed as $205.00 + GST for one year. This annual subscription will automatically renew at $233.00 + GST every 52 weeks (10% off the regular annual price of $259.35). Offer available to new and qualified returning subscribers only. Cancel any time.

To continue reading, please subscribe:

Add Free Press access to your Brandon Sun subscription for only an additional

$1 for the first 4 weeks*

  • Enjoy unlimited reading on winnipegfreepress.com
  • Read the E-Edition, our digital replica newspaper
  • Access News Break, our award-winning app
  • Play interactive puzzles
Start now

*Your next Brandon Sun subscription payment will increase by $1.00 and you will be charged $17.95 plus GST for four weeks. After four weeks, your payment will increase to $24.95 plus GST every four weeks.

Hey there, time traveller!
This article was published 04/09/2024 (713 days ago), so information in it may no longer be current.

How do you know if an artificial intelligence system is so powerful that it poses a security danger and shouldn’t be unleashed without careful oversight?

For regulators trying to put guardrails on AI, it’s mostly about the arithmetic. Specifically, an AI model trained on 10 to the 26th floating-point operations must now be reported to the U.S. government and could soon trigger even stricter requirements in California.

Say what? Well, if you’re counting the zeroes, that’s 100,000,000,000,000,000,000,000,000, or 100 septillion, calculations to train AI systems on huge troves of data.

FILE - President Joe Biden signs an executive on artificial intelligence in the East Room of the White House, Oct. 30, 2023, in Washington. Vice President Kamala Harris looks on at right. (AP Photo/Evan Vucci, File)
FILE - President Joe Biden signs an executive on artificial intelligence in the East Room of the White House, Oct. 30, 2023, in Washington. Vice President Kamala Harris looks on at right. (AP Photo/Evan Vucci, File)

What it signals to some lawmakers and AI safety advocates is a level of computing power that might enable rapidly advancing AI technology to create or proliferate weapons of mass destruction, or conduct catastrophic cyberattacks.

Those who’ve crafted such regulations acknowledge they are an imperfect starting point to distinguish today’s highest-performing generative AI systems — largely made by California-based companies like Anthropic, Google, Meta Platforms and ChatGPT-maker OpenAI — from the next generation that could be even more powerful.

Critics have pounced on the thresholds as arbitrary — an attempt by governments to regulate math. Adding to the confusion is that some rules set a speed-based computing threshold — how many floating-point operations per second, known as flops — while others are based on cumulative number of calculations no matter how long they take.

“Ten to the 26th flops,” said venture capitalist Ben Horowitz on a podcast this summer. “Well, what if that’s the size of the model you need to, like, cure cancer?”

An executive order signed by President Joe Biden last year relies on a 10 to the 26th threshold. So does California’s newly passed AI safety legislation — which Gov. Gavin Newsom has until Sept. 30 to sign into law or veto. California adds a second metric to the equation: regulated AI models must also cost at least $100 million to build.

Following Biden’s footsteps, the European Union’s sweeping AI Act also measures floating-point operations, but sets the bar 10 times lower at 10 to the 25th power. That covers some AI systems already in operation. China’s government has also looked at measuring computing power to determine which AI systems need safeguards.

No publicly available models meet the higher California threshold, though it’s likely that some companies have already started to build them. If so, they’re supposed to be sharing certain details and safety precautions with the U.S. government. Biden employed a Korean War-era law to compel tech companies to alert the U.S. Commerce Department if they’re building such AI models.

AI researchers are still debating how best to evaluate the capabilities of the latest generative AI technology and how it compares to human intelligence. There are tests that judge AI on solving puzzles, logical reasoning or how swiftly and accurately it predicts what text will answer a person’s chatbot query. Those measurements help assess an AI tool’s usefulness for a given task, but there’s no easy way of knowing which one is so widely capable that it poses a danger to humanity.

“This computation, this flop number, by general consensus is sort of the best thing we have along those lines,” said physicist Anthony Aguirre, executive director of the Future of Life Institute, which has advocated for the passage of California’s Senate Bill 1047 and other AI safety rules around the world.

Floating point arithmetic might sound fancy “but it’s really just numbers that are being added or multiplied together,” making it one of the simplest ways to assess an AI model’s capability and risk, Aguirre said.

“Most of what these things are doing is just multiplying big tables of numbers together,” he said. “You can just think of typing in a couple of numbers into your calculator and adding or multiplying them. And that’s what it’s doing — ten trillion times or a hundred trillion times.”

For some tech leaders, however, it’s too simple and hard-coded a metric. There’s “no clear scientific support” for using such metrics as a proxy for risk, argued computer scientist Sara Hooker, who leads AI company Cohere’s nonprofit research division, in a July paper.

“Compute thresholds as currently implemented are shortsighted and likely to fail to mitigate risk,” she wrote.

Venture capitalist Horowitz and his business partner Marc Andreessen, founders of the influential Silicon Valley investment firm Andreessen Horowitz, have attacked the Biden administration as well as California lawmakers for AI regulations they argue could snuff out an emerging AI startup industry.

For Horowitz, putting limits on “how much math you’re allowed to do” reflects a mistaken belief there will only be a handful of big companies making the most capable models and you can put “flaming hoops in front of them and they’ll jump through them and it’s fine.”

In response to the criticism, the sponsor of California’s legislation sent a letter to Andreessen Horowitz this summer defending the bill, including its regulatory thresholds.

Regulating at over 10 to the 26th is “a clear way to exclude from safety testing requirements many models that we know, based on current evidence, lack the ability to cause critical harm,” wrote state Sen. Scott Wiener of San Francisco. Existing publicly released models “have been tested for highly hazardous capabilities and would not be covered by the bill,” Wiener said.

Both Wiener and the Biden executive order treat the metric as a temporary one that could be adjusted later.

Yacine Jernite, who works on policy research at the AI company Hugging Face, said the computing metric emerged in “good faith” ahead of last year’s Biden order but is already starting to grow obsolete. AI developers are doing more with smaller models requiring less computing power, while the potential harms of more widely used AI products won’t trigger California’s proposed scrutiny.

“Some models are going to have a drastically larger impact on society, and those should be held to a higher standard, whereas some others are more exploratory and it might not make sense to have the same kind of process to certify them,” Jernite said.

Aguirre said it makes sense for regulators to be nimble, but he characterizes some opposition to the threshold as an attempt to avoid any regulation of AI systems as they grow more capable.

“This is all happening very fast,” Aguirre said. “I think there’s a legitimate criticism that these thresholds are not capturing exactly what we want them to capture. But I think it’s a poor argument to go from that to, ‘Well, we just shouldn’t do anything and just cross our fingers and hope for the best.’”

Report Error Submit a Tip

More Stories

Free fix is in: repair café aims to keep items out of landfills

Tiago Resko 3 minute read Preview

Free fix is in: repair café aims to keep items out of landfills

Tiago Resko 3 minute read Updated: Yesterday at 12:09 PM CDT

Stephen Kirk wants to give people a chance to fix items they’re not ready to let go of.

The Spence Neighbourhood Association staffer took notice when repair cafés started popping up around the world, and he decided to bring the idea to Winnipeg.

The café invites people to bring broken items like bikes, clothes and small kitchen appliances to the association’s headquarters to get free repairs by a team of skilled volunteers.

“Certainly in our neighbourhood, a lot of people’s motivation is economic, they would rather fix it than have to buy a new one,” said Kirk, the environmental and open spaces coordinator for the Spence Neighbourhood Association.

Read
Updated: Yesterday at 12:09 PM CDT

Puzzles Palace

1 minute read Monday, Jul. 27, 2026

To solve our puzzles, please subscribe with this special offer: |

Liquor authority fines True North $12K after minors with fake ID served at Jets White Out party

Nicole Buffie 5 minute read Preview

Liquor authority fines True North $12K after minors with fake ID served at Jets White Out party

Nicole Buffie 5 minute read Yesterday at 6:17 PM CDT

True North Sports and Entertainment was hit with a $12,000 fine for serving minors with fake IDs during a Winnipeg Jets White Out street party in 2025.

Read
Yesterday at 6:17 PM CDT

Kinew pledges temporary ER in Dauphin by Labour Day

Free Press staff 2 minute read Preview

Kinew pledges temporary ER in Dauphin by Labour Day

Free Press staff 2 minute read Updated: 9:49 AM CDT

Dauphin’s 8,400 residents, ravaged by floodwaters from unprecedented rainfall in June that has left the city without its hospital for the entire summer, will have a temporary but functioning ER shortly, Premier Wab Kinew promised Monday.

Read
Updated: 9:49 AM CDT

It’s back and it’s better than ever.

National non-profit ogranization motionball will be returning for its annual event — the 2026 Marathon of Sports Winnipeg — at the Dakota Community Centre Sept. 19. This fundraising event will help support the Special Olympics Manitoba and Canada foundation in hopes of reaching their goal of $100,000.

Local Special Olympic athletes, including the province’s own Alec Baldwin, as well as registered participants will join together to play a variety of team sports such as basketball, kinball and bocce ball.

Two-time cancer survivor trades blows with angry Lake Winnipeg in epic 12-hour fundraising swim

Zoe Pierce 3 minute read Preview

Two-time cancer survivor trades blows with angry Lake Winnipeg in epic 12-hour fundraising swim

Zoe Pierce 3 minute read Updated: Yesterday at 2:46 PM CDT

After 12 hours of wind, waves and a painful shoulder, Jonathon Fenton has finished what he started, making it across Lake Winnipeg.

The 62-year-old two-time cancer survivor completed the roughly 26-kilometre swim Sunday, a year after shoulder problems forced him to stop 16 kilometres into his first attempt.

The swim was the second year of Fenton’s Jonny’s Big Swim fundraiser, which supports the Alberta Cancer Foundation, CancerCare Manitoba Foundation and Health Sciences Centre Foundation.

He raised roughly $48,000 this year, bringing the two-year total to about $90,000.

Read
Updated: Yesterday at 2:46 PM CDT