OpenAI claims New York Times copyright lawsuit is without merit

10:56 AM PST • January 8, 2024

Image Credits: Bryce Durbin / TechCrunch

In late December, The New York Times sued OpenAI and its close collaborator and investor, Microsoft, for allegedly violating copyright law by training generative AI models on the Times’ content. Today, OpenAI gave a public response, claiming — unsurprisingly — that the Times’ lawsuit is meritless.

In a letter published this afternoon on OpenAI’s official blog, the company reiterates its view that training AI models using publicly available data from the web — including articles like the Times’ — is fair use. In other words, in creating generative AI systems like GPT-4 and DALL-E 3, which “learn” from billions of examples of artwork, ebooks, essays and more to generate human-like text and images, OpenAI believes that it isn’t required to license or otherwise pay for the examples — even if it makes money from those models.

“We view this principle as fair to creators, necessary for innovators and critical for U.S. competitiveness,” OpenAI writes.

OpenAI also addresses in its letter regurgitation, the phenomenon where generative AI models spit out training data verbatim (or near-verbatim) when prompted in a certain way — for example, generating a photo that’s identical to one taken by a famous photographer. OpenAI makes the case that regurgitation is less likely to occur with training data from a single source (e.g., The New York Times) and places the onus on users to “act responsibly” and avoid intentionally prompting its models to regurgitate.

“Interestingly, the regurgitations The New York Times [cites in its lawsuit] appear to be from years-old articles that have proliferated on multiple third-party websites,” OpenAI writes. “It seems they intentionally manipulated prompts, often including lengthy excerpts of articles, in order to get our model to regurgitate. Even when using such prompts, our models don’t typically behave the way The New York Times insinuates, which suggests they either instructed the model to regurgitate or cherry-picked their examples from many attempts.”

OpenAI’s response comes as the copyright debate around generative AI reaches a fever pitch.

In a piece published this week in IEEE Spectrum, noted AI critic Gary Marcus and Reid Southen, a visual effects artist, show how AI systems, including DALL-E 3, regurgitate data even when not specifically prompted to do so — making OpenAI’s claims to the contrary less credible. Marcus and Southen, in fact, make reference to The New York Times lawsuit in their piece, noting that the Times was able to elicit “plagiaristic” responses from OpenAI’s models simply by giving the first few words from a Times story.

The Times is only the latest copyright holder to sue OpenAI over what it believes is a clear violation of IP laws.

Actress Sarah Silverman joined a pair of lawsuits in July that accuse Meta and OpenAI of having “ingested” Silverman’s memoir to train their AI models. In a separate suit, thousands of novelists, including Jonathan Franzen and John Grisham, claim OpenAI sourced their work as training data without their permission or knowledge. And several programmers have an ongoing case against Microsoft, OpenAI and GitHub over Copilot, an AI-powered code-generating tool, which the plaintiffs say was developed using their IP-protected code.

More TechCrunch

Poshmark’s ‘Promoted Closet’ tool lets sellers boost all their listings at once

Lauren Forristal

38 mins ago

Poshmark, the social commerce site that lets people buy and sell new and used items to each other, launched a paid marketing tool on Thursday, giving sellers the ability to…

Poshmark’s ‘Promoted Closet’ tool lets sellers boost all their listings at once

Google adds Gemini to its Education suite

Ivan Mehta

38 mins ago

Google is launching a Gemini add-on for educational institutes through Google Workspace.

Google adds Gemini to its Education suite

YC-backed Recall.ai gets $10M Series A to help companies use virtual meeting data

Kate Park

38 mins ago

More money for the generative AI boom: Y Combinator-backed developer infrastructure startup Recall.ai announced Thursday it’s raised a $10 million Series A funding round, bringing its total raised to over $12M.…

YC-backed Recall.ai gets $10M Series A to help companies use virtual meeting data

Enterprise

Colab’s collaborative tools for engineers line up $21M in new funding

Kyle Wiggers

38 mins ago

Engineers Adam Keating and Jeremy Andrews were tired of using spreadsheets and screenshots to collab with teammates — so they launched a startup, Colab, to build a better way. The…

Colab’s collaborative tools for engineers line up $21M in new funding

Apps

Reddit reintroduces its awards system

Ivan Mehta

49 mins ago

Reddit announced on Wednesday that it is reintroducing its awards system after shutting down the program last year. The company said that most of the mechanisms related to awards will…

Enterprise

Sigma is building a suite of collaborative data analytics tools

Kyle Wiggers

1 hour ago

Sigma Computing, a startup building a range of data analytics and business intelligence tools, has raised $200 million in a fresh VC round.

Sigma is building a suite of collaborative data analytics tools

Government & Policy

EU ‘closely’ monitoring X in wake of Fico shooting as DSA disinfo probe rumbles on

Natasha Lomas

1 hour ago

European Union enforcers of the bloc’s online governance regime, the Digital Services Act (DSA), said Thursday they’re closely monitoring disinformation campaigns on the Elon Musk-owned social network X (formerly Twitter)…

EU ‘closely’ monitoring X in wake of Fico shooting as DSA disinfo probe rumbles on

Climate

Spoor uses AI to save birds from wind turbines

Rebecca Szkutak

2 hours ago

Wind is the largest source of renewable energy in the U.S., according to the U.S. Energy Information Administration, but wind farms come with an environmental cost as wind turbines can…

Spoor uses AI to save birds from wind turbines

Fintech

Cannabis and gaming payments startup Aeropay is now offering an alternative to Mastercard and Visa

Christine Hall

3 hours ago

The key to taking on legacy players in the financial technology industry may be to go where they have not gone before. That’s what Chicago-based Aeropay is doing. The provider…

Cannabis and gaming payments startup Aeropay is now offering an alternative to Mastercard and Visa

Government & Policy

EU opens child safety probes of Facebook and Instagram, citing addictive design concerns

Natasha Lomas

3 hours ago

Facebook and Instagram are under formal investigation in the European Union over child protection concerns, the Commission announced Thursday. The proceedings follow a raft of requests for information to parent…

EU opens child safety probes of Facebook and Instagram, citing addictive design concerns

Climate

Forget EVs: Why Bedrock Materials is targeting gas-powered cars for its first sodium-ion batteries

Tim De Chant

3 hours ago

Bedrock Materials is developing a new type of sodium-ion battery, which promises to be dramatically cheaper than lithium-ion.

Forget EVs: Why Bedrock Materials is targeting gas-powered cars for its first sodium-ion batteries

Security

Thoma Bravo’s LogRhythm merges with Exabeam in more cybersecurity consolidation

Paul Sawers

3 hours ago

Private equity giant Thoma Bravo has announced that its security information and event management (SIEM) company LogRhythm will be merging with Exabeam, a rival cybersecurity company backed by the likes…

Thoma Bravo’s LogRhythm merges with Exabeam in more cybersecurity consolidation

Temu accused of breaching EU’s DSA in bundle of consumer complaints

Natasha Lomas

9 hours ago

Consumer protection groups around the European Union have filed coordinated complaints against Temu, accusing the Chinese-owned ultra low-cost e-commerce platform of a raft of breaches related to the bloc’s Digital…

Temu accused of breaching EU’s DSA in bundle of consumer complaints

Hardware

Google I/O 2024: Here’s everything Google just announced

Christine Hall

16 hours ago

Here are quick hits of the biggest news from the keynote as they are announced.

Google I/O 2024: Here’s everything Google just announced

Government & Policy

Senate study proposes ‘at least’ $32B yearly for AI programs

Devin Coldewey

17 hours ago

The AI industry moves faster than the rest of the technology sector, which means it outpaces the federal government by several orders of magnitude.

Senate study proposes ‘at least’ $32B yearly for AI programs

Security

FBI seizes hacking forum BreachForums — again

Lorenzo Franceschi-Bicchierai

18 hours ago

The FBI along with a coalition of international law enforcement agencies seized the notorious cybercrime forum BreachForums on Wednesday. For years, BreachForums has been a popular English-language forum for hackers…

FBI seizes hacking forum BreachForums — again

Media & Entertainment

Netflix to take on Google and Amazon by building its own ad server

Lauren Forristal

18 hours ago

The announcement signifies a significant shake-up in the streaming giant’s advertising approach.

Netflix to take on Google and Amazon by building its own ad server

Enterprise

Matt Garman taking over as CEO with AWS at crossroads

Ron Miller

19 hours ago

It’s tough to say that a $100 billion business finds itself at a critical juncture, but that’s the case with Amazon Web Services, the cloud arm of Amazon, and the…

Matt Garman taking over as CEO with AWS at crossroads

Google still hasn’t fixed Gemini’s biased image generator

Kyle Wiggers

19 hours ago

Back in February, Google paused its AI-powered chatbot Gemini’s ability to generate images of people after users complained of historical inaccuracies. Told to depict “a Roman legion,” for example, Gemini would show…

Google still hasn’t fixed Gemini’s biased image generator

Privacy

Google’s call-scanning AI could dial up censorship by default, privacy experts warn

Natasha Lomas

20 hours ago

A feature Google demoed at its I/O confab yesterday, using its generative AI technology to scan voice calls in real time for conversational patterns associated with financial scams, has sent…

Google’s call-scanning AI could dial up censorship by default, privacy experts warn

The top AI announcements from Google I/O

Kyle Wiggers

20 hours ago

Google’s going all in on AI — and it wants you to know it. During the company’s keynote at its I/O developer conference on Tuesday, Google mentioned “AI” more than…

The top AI announcements from Google I/O

Transportation

Uber has a new way to solve the concert traffic problem

Rebecca Bellan

20 hours ago

Uber is taking a shuttle product it developed for commuters in India and Egypt and converting it for an American audience. The ride-hail and delivery giant announced Wednesday at its…

Uber has a new way to solve the concert traffic problem

Google takes aim at Android malware with an AI-powered live threat detection service

Sarah Perez

21 hours ago

Google is preparing to launch a new system to help address the problem of malware on Android. Its new live threat detection service leverages Google Play Protect’s on-device AI to…

Apps

Google Maps is getting geospatial AR content later this year

Aisha Malik

21 hours ago

Users will be able to access the AR content by first searching for a location in Google Maps.

Google Maps is getting geospatial AR content later this year

Climate

Quilt heat pump sports sleek design from veterans of Apple, Tesla and Nest

Tim De Chant

21 hours ago

The heat pump startup unveiled its first products and revealed details about performance, pricing and availability.

Quilt heat pump sports sleek design from veterans of Apple, Tesla and Nest

Apps

Google’s new Private Space feature is like Incognito Mode for Android

Brian Heater

21 hours ago

The space is available from the launcher and can be locked as a second layer of authentication.

Google’s new Private Space feature is like Incognito Mode for Android

Media & Entertainment

Google TV to launch AI-generated movie descriptions

Lauren Forristal

21 hours ago

Gemini, the company’s family of generative AI models, will enhance the smart TV operating system so it can generate descriptions for movies and TV shows.

Google TV to launch AI-generated movie descriptions

Hardware

Android’s new Theft Detection Lock helps deter smartphone snatch and grabs

Brian Heater

21 hours ago

When triggered, the AI-powered feature will automatically lock the device down.

Android’s new Theft Detection Lock helps deter smartphone snatch and grabs

Security

Google adds live threat detection and screen-sharing protection to Android

Ivan Mehta

21 hours ago

The company said it is increasing the on-device capability of its Google Play Protect system to detect fraudulent apps trying to breach sensitive permissions.

Google adds live threat detection and screen-sharing protection to Android

Apps

Wear OS 5 hits developer preview, offering better battery life

Sarah Perez

21 hours ago

This latest release, one of many announcements from the Google I/O 2024 developer conference, focuses on improved battery life and other performance improvements, like more efficient workout tracking.

OpenAI claims New York Times copyright lawsuit is without merit

More TechCrunch

Get the industry’s biggest tech news

TechCrunch Daily News

Startups Weekly

TechCrunch Fintech

TechCrunch Mobility

Tags