How do you know when AI is powerful enough to be dangerous? Regulators try to do the math

Advertisement

Advertise with us

How do you know if an artificial intelligence system is so powerful that it poses a security danger and shouldn’t be unleashed without careful oversight?

Read this article for free:


or

Already have an account? Log in here »

To continue reading, please subscribe:

Subscribe and receive a limited-edition Free Press branded hat or tote.

Digital Subscription

One year of digital access for only $205*

  • Enjoy unlimited reading on winnipegfreepress.com
  • Read the E-Edition, our digital replica newspaper
  • Access News Break, our award-winning app
  • Play interactive puzzles

*First annual payment billed as $205.00 + GST for one year. This annual subscription will automatically renew at $233.00 + GST every 52 weeks (10% off the regular annual price of $259.35). Offer available to new and qualified returning subscribers only. Cancel any time.

To continue reading, please subscribe:

Add Free Press access to your Brandon Sun subscription for only an additional

$1 for the first 4 weeks*

  • Enjoy unlimited reading on winnipegfreepress.com
  • Read the E-Edition, our digital replica newspaper
  • Access News Break, our award-winning app
  • Play interactive puzzles
Start now

*Your next Brandon Sun subscription payment will increase by $1.00 and you will be charged $17.95 plus GST for four weeks. After four weeks, your payment will increase to $24.95 plus GST every four weeks.

Hey there, time traveller!
This article was published 04/09/2024 (710 days ago), so information in it may no longer be current.

How do you know if an artificial intelligence system is so powerful that it poses a security danger and shouldn’t be unleashed without careful oversight?

For regulators trying to put guardrails on AI, it’s mostly about the arithmetic. Specifically, an AI model trained on 10 to the 26th floating-point operations must now be reported to the U.S. government and could soon trigger even stricter requirements in California.

Say what? Well, if you’re counting the zeroes, that’s 100,000,000,000,000,000,000,000,000, or 100 septillion, calculations to train AI systems on huge troves of data.

FILE - President Joe Biden signs an executive on artificial intelligence in the East Room of the White House, Oct. 30, 2023, in Washington. Vice President Kamala Harris looks on at right. (AP Photo/Evan Vucci, File)
FILE - President Joe Biden signs an executive on artificial intelligence in the East Room of the White House, Oct. 30, 2023, in Washington. Vice President Kamala Harris looks on at right. (AP Photo/Evan Vucci, File)

What it signals to some lawmakers and AI safety advocates is a level of computing power that might enable rapidly advancing AI technology to create or proliferate weapons of mass destruction, or conduct catastrophic cyberattacks.

Those who’ve crafted such regulations acknowledge they are an imperfect starting point to distinguish today’s highest-performing generative AI systems — largely made by California-based companies like Anthropic, Google, Meta Platforms and ChatGPT-maker OpenAI — from the next generation that could be even more powerful.

Critics have pounced on the thresholds as arbitrary — an attempt by governments to regulate math. Adding to the confusion is that some rules set a speed-based computing threshold — how many floating-point operations per second, known as flops — while others are based on cumulative number of calculations no matter how long they take.

“Ten to the 26th flops,” said venture capitalist Ben Horowitz on a podcast this summer. “Well, what if that’s the size of the model you need to, like, cure cancer?”

An executive order signed by President Joe Biden last year relies on a 10 to the 26th threshold. So does California’s newly passed AI safety legislation — which Gov. Gavin Newsom has until Sept. 30 to sign into law or veto. California adds a second metric to the equation: regulated AI models must also cost at least $100 million to build.

Following Biden’s footsteps, the European Union’s sweeping AI Act also measures floating-point operations, but sets the bar 10 times lower at 10 to the 25th power. That covers some AI systems already in operation. China’s government has also looked at measuring computing power to determine which AI systems need safeguards.

No publicly available models meet the higher California threshold, though it’s likely that some companies have already started to build them. If so, they’re supposed to be sharing certain details and safety precautions with the U.S. government. Biden employed a Korean War-era law to compel tech companies to alert the U.S. Commerce Department if they’re building such AI models.

AI researchers are still debating how best to evaluate the capabilities of the latest generative AI technology and how it compares to human intelligence. There are tests that judge AI on solving puzzles, logical reasoning or how swiftly and accurately it predicts what text will answer a person’s chatbot query. Those measurements help assess an AI tool’s usefulness for a given task, but there’s no easy way of knowing which one is so widely capable that it poses a danger to humanity.

“This computation, this flop number, by general consensus is sort of the best thing we have along those lines,” said physicist Anthony Aguirre, executive director of the Future of Life Institute, which has advocated for the passage of California’s Senate Bill 1047 and other AI safety rules around the world.

Floating point arithmetic might sound fancy “but it’s really just numbers that are being added or multiplied together,” making it one of the simplest ways to assess an AI model’s capability and risk, Aguirre said.

“Most of what these things are doing is just multiplying big tables of numbers together,” he said. “You can just think of typing in a couple of numbers into your calculator and adding or multiplying them. And that’s what it’s doing — ten trillion times or a hundred trillion times.”

For some tech leaders, however, it’s too simple and hard-coded a metric. There’s “no clear scientific support” for using such metrics as a proxy for risk, argued computer scientist Sara Hooker, who leads AI company Cohere’s nonprofit research division, in a July paper.

“Compute thresholds as currently implemented are shortsighted and likely to fail to mitigate risk,” she wrote.

Venture capitalist Horowitz and his business partner Marc Andreessen, founders of the influential Silicon Valley investment firm Andreessen Horowitz, have attacked the Biden administration as well as California lawmakers for AI regulations they argue could snuff out an emerging AI startup industry.

For Horowitz, putting limits on “how much math you’re allowed to do” reflects a mistaken belief there will only be a handful of big companies making the most capable models and you can put “flaming hoops in front of them and they’ll jump through them and it’s fine.”

In response to the criticism, the sponsor of California’s legislation sent a letter to Andreessen Horowitz this summer defending the bill, including its regulatory thresholds.

Regulating at over 10 to the 26th is “a clear way to exclude from safety testing requirements many models that we know, based on current evidence, lack the ability to cause critical harm,” wrote state Sen. Scott Wiener of San Francisco. Existing publicly released models “have been tested for highly hazardous capabilities and would not be covered by the bill,” Wiener said.

Both Wiener and the Biden executive order treat the metric as a temporary one that could be adjusted later.

Yacine Jernite, who works on policy research at the AI company Hugging Face, said the computing metric emerged in “good faith” ahead of last year’s Biden order but is already starting to grow obsolete. AI developers are doing more with smaller models requiring less computing power, while the potential harms of more widely used AI products won’t trigger California’s proposed scrutiny.

“Some models are going to have a drastically larger impact on society, and those should be held to a higher standard, whereas some others are more exploratory and it might not make sense to have the same kind of process to certify them,” Jernite said.

Aguirre said it makes sense for regulators to be nimble, but he characterizes some opposition to the threshold as an attempt to avoid any regulation of AI systems as they grow more capable.

“This is all happening very fast,” Aguirre said. “I think there’s a legitimate criticism that these thresholds are not capturing exactly what we want them to capture. But I think it’s a poor argument to go from that to, ‘Well, we just shouldn’t do anything and just cross our fingers and hope for the best.’”

Report Error Submit a Tip

More Stories

WRHA directs three care homes to accommodate surge of hospital patients

Malak Abas 5 minute read Preview

WRHA directs three care homes to accommodate surge of hospital patients

Malak Abas 5 minute read Updated: 5:01 PM CDT

The Winnipeg Regional Health Authority has directed three long-term care centres to make space for hospital patients, citing “exceptionally high demand for care across the health-care system.”

Read
Updated: 5:01 PM CDT

Icy relationship with U.S. beginning to thaw at border, data shows; American visits to Manitoba hit five-year high

Nicole Buffie 5 minute read Preview

Icy relationship with U.S. beginning to thaw at border, data shows; American visits to Manitoba hit five-year high

Nicole Buffie 5 minute read Yesterday at 5:32 PM CDT

As Canadians begin to lower their elbows and start to cross the U.S. border again, Americans are coming to Manitoba in numbers not seen in years.

Read
Yesterday at 5:32 PM CDT

Three security guards charged with assault after Sea Bears star injured outside nightclub

Nicole Buffie 3 minute read Preview

Three security guards charged with assault after Sea Bears star injured outside nightclub

Nicole Buffie 3 minute read Updated: Yesterday at 12:50 PM CDT

Three security guards are facing assault charges, accused of injuring a Winnipeg Sea Bears player at a nightclub earlier this month.

Read
Updated: Yesterday at 12:50 PM CDT

Development turning long-vacant gas station in Selkirk into commercial hub

By Tiago Resko 2 minute read Preview

Development turning long-vacant gas station in Selkirk into commercial hub

By Tiago Resko 2 minute read Yesterday at 3:59 PM CDT

A plot of land that has sat vacant in Selkirk for 30 years is being redeveloped to host new businesses in the heart of the city.

The former gas station site on Main Street — between Rosser and Pacific avenues — will be turned into a commercial hub that will be home to six local businesses, and will be a big stepping stone for economic growth in Selkirk, Mayor Larry Johannson said Thursday.

“For decades this site has been a reminder of what once was. Today, it represents what is possible when investors see the potential of our community and choose to be a part of its future,” he said in a news release announcing the project and Thursday’s groundbreaking.

“This development reflects confidence in Selkirk, our economy and the direction we are heading.”

Read
Yesterday at 3:59 PM CDT

Proposed Transcona development — which could add 20,000 people — gets thumbs down from residents

Joyanne Pursaga 5 minute read Preview

Proposed Transcona development — which could add 20,000 people — gets thumbs down from residents

Joyanne Pursaga 5 minute read 6:15 PM CDT

Dozens of Winnipeggers fear plans to make way for up to 20,000 more people to live in south Transcona would dramatically alter their neighbourhood and tie hefty fees for new services to their property titles.

If city council adopts the proposed South Transcona secondary plan, all future development applications would need to conform to its goals. The affected area is located south of Dugald Road, north of St. Boniface Road, east of Plessis Road and west of Murdock Road.

Homeowner Virginia Page Jähne fears the plan could add a massive tab to the title of her 4.3-acre property to cover services required for the development.

“My main concern is that they are wanting to put $1.4 million onto my title to pay for this. Now, they say that I will not have to pay it unless I subdivide. But if a person comes by and wants to buy from me … and they see that $1.4 million … what’s the first thing (they’re) going to do?” said Page Jähne, in an interview.

Read
6:15 PM CDT

Manitoba prepares project pitches for investment summit

Gabrielle Piché 6 minute read Preview

Manitoba prepares project pitches for investment summit

Gabrielle Piché 6 minute read 6:17 PM CDT

Powerhouse investors may hear pitches on Manitoba mines, silica sand projects and a massive greenhouse during Canada’s first major investors summit.

The Manitoba government is finalizing its list of projects to pitch at the Canada Investment Summit next month.

“The conversations that’ll be had… will be about investments in the billions of dollars,” Premier Wab Kinew said Friday.

“It’ll be one step among many that we take to build a bright future in Manitoba.”

Read
6:17 PM CDT