AI

Ghost, now OpenAI-backed, claims LLMs will overcome self-driving setbacks — but experts are skeptical

Comment

Ghost Autonomy
Image Credits: Ghost Autonomy

It’s not hyperbolic to say that the self-driving car industry is facing a reckoning.

Just this week, Cruise recalled its entire fleet of autonomous cars after a grisly accident involving a pedestrian that led the California DMV to suspend the company from operating driverless robotaxis in the state. Meanwhile, activists in San Francisco have taken to the streets — literally — to immobilize driverless cars as form of protest against the city being used as a testing ground for the emerging technology.

But one startup says it holds the key to safer self-driving technology — and thinks that this key will convince the naysayers.

Ghost Autonomy, a company building autonomous driving software for automaker partners, this week announced that it plans to begin exploring the applications of multimodal large language models (LLMs) — AI models that can understand text as well as images — in self-driving. To realize this, Ghost has partnered with OpenAI through the OpenAI Startup Fund to gain early access to OpenAI systems and Azure resources from Microsoft, OpenAI’s close collaborator, plus a $5 million investment.

“LLMs offer a new way to understand ‘the long tail,’ adding reasoning to complex scenes where current models fall short,” Ghost co-founder and CEO John Hayes told TechCrunch in an email interview. “The use cases for LLM-based analysis in autonomy will only grow as LLMs get faster and more capable.”

But how, exactly, is Ghost applying AI models designed to explain images and generate text to controlling autonomous cars? According to Hayes, Ghost is piloting software that relies on multimodal models to “do higher complexity scene interpretation.” suggesting road decisions (e.g. “move to the right lane”) to car-controlling hardware based on pictures of road scenes from car-mounted cameras.

“At Ghost, we’ll be working to fine-tune existing models and training our own models to maximize reliability and performance on the road,” Hayes said. “For example, construction zones have unusual components that can be difficult for simpler models to navigate — temporary lanes, flagmen holding signs that change, and complex negotiation with other road users. LLMs have shown to be able to process all of these variables in concert with human-like levels of reasoning.”

The experts I spoke with are skeptical, however.

“[Ghost is] using ‘LLM’ as a marketing buzzword,” Os Keyes, a Ph.D. candidate at the University of Washington focusing on law and data ethics, told TechCrunch via email. “Basically, if you take this pitch and replaced LLM with ‘blockchain’ and sent it back to 2016, it would be just as plausible — and just as obviously a boondoggle.”

Keyes posits that LLMs are simply the wrong tool for self-driving. They weren’t designed or trained for this purpose, he asserts, and may even be a less efficient way of solving some of the outstanding challenges in vehicular autonomy.

“It’s sort of like hearing your neighbor has been using a sheaf of treasury notes to hold a table up,” Keyes said. “You could do it that way, and it’s certainly fancier than the alternative, but… why?”

Mike Cook, a senior lecturer at King’s College London whose research focuses on computational creativity, agrees with Keyes’ overall assessment. He notes that multimodal models themselves are far from a solved science; indeed, OpenAI’s flagship model invents facts and makes basic mistakes that humans wouldn’t, like copying down text incorrectly and getting colors wrong.

“I don’t believe there’s any such thing as a silver bullet in computer science,” Cook said. “There’s simply no reason to put LLMs at the center of something as dangerous and complex as driving a car. Researchers around the world are already struggling to find ways to validate and prove the safety of LLMs for fairly ordinary tasks like answering essay questions, and the idea that we should be applying this often unpredictable and unstable technology to autonomous driving is premature at best — and misguided at worst.”

But Hayes and OpenAI won’t be dissuaded.

In a press release, Brad Lightcap, OpenAI’s COO and manager of the OpenAI Startup Fund, is quoted as saying that multimodal models “have the potential to expand the applicability of LLMs to many new use cases,” including autonomy and automotive. He adds: “With the ability to understand and draw conclusions by combining video, images and sounds, multimodal models may create a new way to understand scenes and navigate complex or unusual environments.”

TechCrunch emailed questions to Lightcap via OpenAI’s press relations but hadn’t heard back as of publication time.

As for Hayes, he says argues that LLMs could allow autonomous driving systems to “reason about driving scenes holistically” and “utilize broad-based world knowledge” to “navigate complex and unusual situations” —  even situations they hadn’t seen before. He claims that Ghost is actively testing multimodal model-driving decision making via its development fleet and working with automakers to “jointly validate” and integrate new large models into Ghost’s autonomy stack.

“No doubt the current models are not quite ready for commercial use in cars,” Hayes said. “There’s still a lot of work to do to improve their reliability and performance. But this is exactly why there’s a market for application-specific companies doing R&D on these general models. Companies like ours with lots of training data and a deep understanding of the application will dramatically improve upon the existing general models. The models themselves will also improve …. Ultimately, autonomous driving will require a complete system to deliver safety, with many different model types and functions. [Multimodal models] are just one tool to help make that happen.”

That’s promising a lot with unproven tech. Can Ghost deliver? Given companies as well-financed and well-resourced as Cruise and Waymo are experiencing major setbacks many years into testing self-driving vehicles on the road, I’m not so sure.

More TechCrunch

Hello and welcome back to TechCrunch Space. Happy belated Mother’s Day! Want to reach out with a tip? Email Aria at aria.techcrunch@gmail.com or send me a message on Signal at…

Apple devoted a full event to iPad last Tuesday, roughly a month out from WWDC. From the invite artwork to the polarizing ad spot, Apple was clear — the event…

Apple iPad Pro M4 vs. iPad Air M2: Reviewing which is right for most

Terri Burns, a former partner at GV, is venturing into a new chapter of her career by launching her own venture firm called Type Capital. 

GV’s youngest partner has launched her own firm

The decision to go monochrome was probably a smart one, considering the candy-colored alternatives that seem to want to dazzle and comfort you.

ChatGPT’s new face is a black hole

Apple and Google announced on Monday that iPhone and Android users will start seeing alerts when it’s possible that an unknown Bluetooth device is being used to track them. The…

Apple and Google agree on standard to alert people when unknown Bluetooth devices may be tracking them

The company is describing the event as “a chance to demo some ChatGPT and GPT-4 updates.”

OpenAI’s ChatGPT announcement: Watch here

A human safety operator will be behind the wheel during this phase of testing, according to the company.

GM’s Cruise ramps up robotaxi testing in Phoenix

OpenAI announced a new flagship generative AI model on Monday that they call GPT-4o — the “o” stands for “omni,” referring to the model’s ability to handle text, speech, and…

OpenAI debuts GPT-4o ‘omni’ model now powering ChatGPT

Featured Article

The women in AI making a difference

As a part of a multi-part series, TechCrunch is highlighting women innovators — from academics to policymakers —in the field of AI.

5 hours ago
The women in AI making a difference

The expansion of Polar Semiconductor’s facility would enable the company to double its U.S. production capacity of sensor and power chips within two years.

White House proposes up to $120 million to help fund Polar Semiconductor’s chip facility expansion

In 2021, Google kicked off work on Project Starline, a corporate-focused teleconferencing platform that uses 3D imaging, cameras and a custom-designed screen to let people converse with someone as if…

Google’s 3D video conferencing platform, Project Starline, is coming in 2025 with help from HP

Over the weekend, Instagram announced it is expanding its creator marketplace to 10 new countries — this marketplace connects brands with creators to foster collaboration. The new regions include South…

Instagram expands its creator marketplace to 10 new countries

You can expect plenty of AI, but probably not a lot of hardware.

Google I/O 2024: What to expect

The keynote kicks off at 10 a.m. PT on Tuesday and will offer glimpses into the latest versions of Android, Wear OS and Android TV.

Google I/O 2024: How to watch

Four-year-old Mexican BNPL startup Aplazo facilitates fractionated payments to offline and online merchants even when the buyer doesn’t have a credit card.

Aplazo is using buy now, pay later as a stepping stone to financial ubiquity in Mexico

We received countless submissions to speak at this year’s Disrupt 2024. After carefully sifting through all the applications, we’ve narrowed it down to 19 session finalists. Now we need your…

Vote for your Disrupt 2024 Audience Choice favs

Co-founder and CEO Bowie Cheung, who previously worked at Uber Eats, said the company now has 200 customers.

Healthy growth helps B2B food e-commerce startup Pepper nab $30 million led by ICONIQ Growth

Booking.com has been designated a gatekeeper under the EU’s DMA, meaning the firm will be regulated under the bloc’s market fairness framework.

Booking.com latest to fall under EU market power rules

Featured Article

‘Got that boomer!’: How cybercriminals steal one-time passcodes for SIM swap attacks and raiding bank accounts

Estate is an invite-only website that has helped hundreds of attackers make thousands of phone calls aimed at stealing account passcodes, according to its leaked database.

10 hours ago
‘Got that boomer!’: How cybercriminals steal one-time passcodes for SIM swap attacks and raiding bank accounts

Squarespace is being taken private in an all-cash deal that values the company on an equity basis at $6.6 billion.

Permira is taking Squarespace private in a $6.9 billion deal

AI-powered tools like OpenAI’s Whisper have enabled many apps to make transcription an integral part of their feature set for personal note-taking, and the space has quickly flourished as a…

Buy Me a Coffee’s founder has built an AI-powered voice note app

Airtel, India’s second-largest telco, is partnering with Google Cloud to develop and deliver cloud and GenAI solutions to Indian businesses.

Google partners with Airtel to offer cloud and GenAI products to Indian businesses

To give AI-focused women academics and others their well-deserved — and overdue — time in the spotlight, TechCrunch has been publishing a series of interviews focused on remarkable women who’ve contributed to…

Women in AI: Rep. Dar’shun Kendrick wants to pass more AI legislation

We took the pulse of emerging fund managers about what it’s been like for them during these post-ZERP, venture-capital-winter years.

A reckoning is coming for emerging venture funds, and that, VCs say, is a good thing

It’s been a busy weekend for union organizing efforts at U.S. Apple stores, with the union at one store voting to authorize a strike, while workers at another store voted…

Workers at a Maryland Apple store authorize strike

Alora Baby is not just aiming to manufacture baby cribs in an environmentally friendly way but is attempting to overhaul the whole lifecycle of a product

Alora Baby aims to push baby gear away from the ‘landfill economy’

Bumble founder and executive chair Whitney Wolfe Herd raised eyebrows this week with her comments about how AI might change the dating experience. During an onstage interview, Bloomberg’s Emily Chang…

Go on, let bots date other bots

Welcome to Week in Review: TechCrunch’s newsletter recapping the week’s biggest news. This week Apple unveiled new iPad models at its Let Loose event, including a new 13-inch display for…

Why Apple’s ‘Crush’ ad is so misguided

The U.K. AI Safety Institute, the U.K.’s recently established AI safety body, has released a toolset designed to “strengthen AI safety” by making it easier for industry, research organizations and…

UK agency releases tools to test AI model safety

AI startup Runway’s second annual AI Film Festival showcased movies that incorporated AI tech in some fashion, from backgrounds to animations.

At the AI Film Festival, humanity triumphed over tech