OpenAI is forming a new team to bring ‘superintelligent’ AI under control

12:07 PM PDT • July 5, 2023

Image Credits: Bryce Durbin / TechCrunch

OpenAI is forming a new team led by Ilya Sutskever, its chief scientist and one of the company’s co-founders, to develop ways to steer and control “superintelligent” AI systems.

In a blog post published today, Sutskever and Jan Leike, a lead on the alignment team at OpenAI, predict that AI with intelligence exceeding that of humans could arrive within the decade. This AI — assuming it does, indeed, arrive eventually — won’t necessarily be benevolent, necessitating research into ways to control and restrict it, Sutskever and Leike say.

“Currently, we don’t have a solution for steering or controlling a potentially superintelligent AI, and preventing it from going rogue,” they write. “Our current techniques for aligning AI, such as reinforcement learning from human feedback, rely on humans’ ability to supervise AI. But humans won’t be able to reliably supervise AI systems much smarter than us.”

To move the needle forward in the area of “superintelligence alignment,” OpenAI is creating a new Superalignment team, led by both Sutskever and Leike, which will have access to 20% of the compute the company has secured to date. Joined by scientists and engineers from OpenAI’s previous alignment division as well as researchers from other orgs across the company, the team will aim to solve the core technical challenges of controlling superintelligent AI over the next four years.

How? By building what Sutskever and Leike describe as a “human-level automated alignment researcher.” The high-level goal is to train AI systems using human feedback, train AI to assist in evaluating other AI systems and ultimately to build AI that can do alignment research. (Here, “alignment research” refers to ensuring AI systems achieve desired outcomes or don’t go off the rails.)

It’s OpenAI’s hypothesis that AI can make faster and better alignment research progress than humans can.

“As we make progress on this, our AI systems can take over more and more of our alignment work and ultimately conceive, implement, study and develop better alignment techniques than we have now,” Leike and colleagues John Schulman and Jeffrey Wu postulated in a previous blog post. “They will work together with humans to ensure that their own successors are more aligned with humans. . . . Human researchers will focus more and more of their effort on reviewing alignment research done by AI systems instead of generating this research by themselves.”

Of course, no method is foolproof — and Leike, Schulman and Wu acknowledge the many limitations of OpenAI in their post. Using AI for evaluation has the potential to scale up inconsistencies, biases or vulnerabilities in that AI, they say. And it might turn out that the hardest parts of the alignment problem might not be related to engineering at all.

But Sutskever and Leike think it’s worth a go.

“Superintelligence alignment is fundamentally a machine learning problem, and we think great machine learning experts — even if they’re not already working on alignment — will be critical to solving it,” they write. “We plan to share the fruits of this effort broadly and view contributing to alignment and safety of non-OpenAI models as an important part of our work.”

More TechCrunch

Agora raises $34B Series B to keep building the Carta for real estate

Marina Temkin

11 mins ago

Since he was very young, Bar Mor knew that he would inevitably do something with real estate. His family was involved in all types of real estate projects, from ground-up…

Agora raises $34B Series B to keep building the Carta for real estate

Commerce

Poshmark’s ‘Promoted Closet’ tool lets sellers boost all their listings at once

Lauren Forristal

1 hour ago

Poshmark, the social commerce site that lets people buy and sell new and used items to each other, launched a paid marketing tool on Thursday, giving sellers the ability to…

Poshmark’s ‘Promoted Closet’ tool lets sellers boost all their listings at once

Google adds Gemini to its Education suite

Ivan Mehta

1 hour ago

Google is launching a Gemini add-on for educational institutes through Google Workspace.

Google adds Gemini to its Education suite

YC-backed Recall.ai gets $10M Series A to help companies use virtual meeting data

Kate Park

1 hour ago

More money for the generative AI boom: Y Combinator-backed developer infrastructure startup Recall.ai announced Thursday it’s raised a $10 million Series A funding round, bringing its total raised to over $12M.…

YC-backed Recall.ai gets $10M Series A to help companies use virtual meeting data

Enterprise

Colab’s collaborative tools for engineers line up $21M in new funding

Kyle Wiggers

1 hour ago

Engineers Adam Keating and Jeremy Andrews were tired of using spreadsheets and screenshots to collab with teammates — so they launched a startup, Colab, to build a better way. The…

Colab’s collaborative tools for engineers line up $21M in new funding

Apps

Reddit reintroduces its awards system

Ivan Mehta

1 hour ago

Reddit announced on Wednesday that it is reintroducing its awards system after shutting down the program last year. The company said that most of the mechanisms related to awards will…

Enterprise

Sigma is building a suite of collaborative data analytics tools

Kyle Wiggers

2 hours ago

Sigma Computing, a startup building a range of data analytics and business intelligence tools, has raised $200 million in a fresh VC round.

Sigma is building a suite of collaborative data analytics tools

Government & Policy

EU ‘closely’ monitoring X in wake of Fico shooting as DSA disinfo probe rumbles on

Natasha Lomas

2 hours ago

European Union enforcers of the bloc’s online governance regime, the Digital Services Act (DSA), said Thursday they’re closely monitoring disinformation campaigns on the Elon Musk-owned social network X (formerly Twitter)…

EU ‘closely’ monitoring X in wake of Fico shooting as DSA disinfo probe rumbles on

Climate

Spoor uses AI to save birds from wind turbines

Rebecca Szkutak

2 hours ago

Wind is the largest source of renewable energy in the U.S., according to the U.S. Energy Information Administration, but wind farms come with an environmental cost as wind turbines can…

Spoor uses AI to save birds from wind turbines

Fintech

Cannabis and gaming payments startup Aeropay is now offering an alternative to Mastercard and Visa

Christine Hall

3 hours ago

The key to taking on legacy players in the financial technology industry may be to go where they have not gone before. That’s what Chicago-based Aeropay is doing. The provider…

Cannabis and gaming payments startup Aeropay is now offering an alternative to Mastercard and Visa

Government & Policy

EU opens child safety probes of Facebook and Instagram, citing addictive design concerns

Natasha Lomas

4 hours ago

Facebook and Instagram are under formal investigation in the European Union over child protection concerns, the Commission announced Thursday. The proceedings follow a raft of requests for information to parent…

EU opens child safety probes of Facebook and Instagram, citing addictive design concerns

Climate

Forget EVs: Why Bedrock Materials is targeting gas-powered cars for its first sodium-ion batteries

Tim De Chant

4 hours ago

Bedrock Materials is developing a new type of sodium-ion battery, which promises to be dramatically cheaper than lithium-ion.

Forget EVs: Why Bedrock Materials is targeting gas-powered cars for its first sodium-ion batteries

Security

Thoma Bravo’s LogRhythm merges with Exabeam in more cybersecurity consolidation

Paul Sawers

4 hours ago

Private equity giant Thoma Bravo has announced that its security information and event management (SIEM) company LogRhythm will be merging with Exabeam, a rival cybersecurity company backed by the likes…

Thoma Bravo’s LogRhythm merges with Exabeam in more cybersecurity consolidation

Temu accused of breaching EU’s DSA in bundle of consumer complaints

Natasha Lomas

10 hours ago

Consumer protection groups around the European Union have filed coordinated complaints against Temu, accusing the Chinese-owned ultra low-cost e-commerce platform of a raft of breaches related to the bloc’s Digital…

Temu accused of breaching EU’s DSA in bundle of consumer complaints

Hardware

Google I/O 2024: Here’s everything Google just announced

Christine Hall

16 hours ago

Here are quick hits of the biggest news from the keynote as they are announced.

Google I/O 2024: Here’s everything Google just announced

Government & Policy

Senate study proposes ‘at least’ $32B yearly for AI programs

Devin Coldewey

18 hours ago

The AI industry moves faster than the rest of the technology sector, which means it outpaces the federal government by several orders of magnitude.

Senate study proposes ‘at least’ $32B yearly for AI programs

Security

FBI seizes hacking forum BreachForums — again

Lorenzo Franceschi-Bicchierai

18 hours ago

The FBI along with a coalition of international law enforcement agencies seized the notorious cybercrime forum BreachForums on Wednesday. For years, BreachForums has been a popular English-language forum for hackers…

FBI seizes hacking forum BreachForums — again

Media & Entertainment

Netflix to take on Google and Amazon by building its own ad server

Lauren Forristal

19 hours ago

The announcement signifies a significant shake-up in the streaming giant’s advertising approach.

Netflix to take on Google and Amazon by building its own ad server

Enterprise

Matt Garman taking over as CEO with AWS at crossroads

Ron Miller

19 hours ago

It’s tough to say that a $100 billion business finds itself at a critical juncture, but that’s the case with Amazon Web Services, the cloud arm of Amazon, and the…

Matt Garman taking over as CEO with AWS at crossroads

Google still hasn’t fixed Gemini’s biased image generator

Kyle Wiggers

20 hours ago

Back in February, Google paused its AI-powered chatbot Gemini’s ability to generate images of people after users complained of historical inaccuracies. Told to depict “a Roman legion,” for example, Gemini would show…

Google still hasn’t fixed Gemini’s biased image generator

Privacy

Google’s call-scanning AI could dial up censorship by default, privacy experts warn

Natasha Lomas

21 hours ago

A feature Google demoed at its I/O confab yesterday, using its generative AI technology to scan voice calls in real time for conversational patterns associated with financial scams, has sent…

Google’s call-scanning AI could dial up censorship by default, privacy experts warn

The top AI announcements from Google I/O

Kyle Wiggers

21 hours ago

Google’s going all in on AI — and it wants you to know it. During the company’s keynote at its I/O developer conference on Tuesday, Google mentioned “AI” more than…

The top AI announcements from Google I/O

Transportation

Uber has a new way to solve the concert traffic problem

Rebecca Bellan

21 hours ago

Uber is taking a shuttle product it developed for commuters in India and Egypt and converting it for an American audience. The ride-hail and delivery giant announced Wednesday at its…

Uber has a new way to solve the concert traffic problem

Google takes aim at Android malware with an AI-powered live threat detection service

Sarah Perez

21 hours ago

Google is preparing to launch a new system to help address the problem of malware on Android. Its new live threat detection service leverages Google Play Protect’s on-device AI to…

Apps

Google Maps is getting geospatial AR content later this year

Aisha Malik

21 hours ago

Users will be able to access the AR content by first searching for a location in Google Maps.

Google Maps is getting geospatial AR content later this year

Climate

Quilt heat pump sports sleek design from veterans of Apple, Tesla and Nest

Tim De Chant

21 hours ago

The heat pump startup unveiled its first products and revealed details about performance, pricing and availability.

Quilt heat pump sports sleek design from veterans of Apple, Tesla and Nest

Apps

Google’s new Private Space feature is like Incognito Mode for Android

Brian Heater

21 hours ago

The space is available from the launcher and can be locked as a second layer of authentication.

Google’s new Private Space feature is like Incognito Mode for Android

Media & Entertainment

Google TV to launch AI-generated movie descriptions

Lauren Forristal

21 hours ago

Gemini, the company’s family of generative AI models, will enhance the smart TV operating system so it can generate descriptions for movies and TV shows.

Google TV to launch AI-generated movie descriptions

Hardware

Android’s new Theft Detection Lock helps deter smartphone snatch and grabs

Brian Heater

21 hours ago

When triggered, the AI-powered feature will automatically lock the device down.

Android’s new Theft Detection Lock helps deter smartphone snatch and grabs

Security

Google adds live threat detection and screen-sharing protection to Android

Ivan Mehta

21 hours ago

The company said it is increasing the on-device capability of its Google Play Protect system to detect fraudulent apps trying to breach sensitive permissions.

OpenAI is forming a new team to bring ‘superintelligent’ AI under control

More TechCrunch

Get the industry’s biggest tech news

TechCrunch Daily News

Startups Weekly

TechCrunch Fintech

TechCrunch Mobility

Tags