AI

OpenAI’s deals with publishers could spell trouble for rivals

Comment

Illustration depicting OpenAI's logo over a flower
Image Credits: Bryce Durbin/TechCrunch

OpenAI’s legal battle with The New York Times over data to train its AI models might still be brewing. But OpenAI’s forging ahead on deals with other publishers, including some of France’s and Spain’s largest news publishers.

OpenAI on Wednesday announced that it signed contracts with Le Monde and Prisa Media to bring French and Spanish news content to OpenAI’s ChatGPT chatbot. In a blog post, OpenAI said that the partnership will put the organizations’ current events coverage — from brands including El País, Cinco Días, As and El Huffpost — in front of ChatGPT users where it makes sense, as well as contribute to OpenAI’s ever-expanding volume of training data.

OpenAI writes:

Over the coming months, ChatGPT users will be able to interact with relevant news content from these publishers through select summaries with attribution and enhanced links to the original articles, giving users the ability to access additional information or related articles from their news sites … We are continually making improvements to ChatGPT and are supporting the essential role of the news industry in delivering real-time, authoritative information to users.

So, OpenAI’s revealed licensing deals with a handful of content providers at this point. Now felt like a good opportunity to take stock:

  • Stock media library Shutterstock (for images, videos and music training data)
  • The Associated Press
  • Axel Springer (owner of Politico and Business Insider, among others)
  • Le Monde
  • Prisa Media

How much is OpenAI paying each? Well, it’s not saying — at least not publicly. But we can estimate.

The Information reported in January that OpenAI was offering publishers between $1 million and $5 million a year to access archives to train its GenAI models. That doesn’t tell us much about the Shutterstock partnership. But on the article licensing front — assuming The Information’s reporting is accurate and those figures haven’t changed since then — OpenAI’s shelling out between $4 million and $20 million a year for news.

That might be pennies to OpenAI, whose war chest sits at over $11 billion and whose annualized revenue recently topped $2 billion (per Financial Times). But as Hunter Walk, a partner at Homebrew and the co-founder of Screendoor, recently mused, it’s substantial enough to potentially edge out AI rivals also pursuing licensing agreements.

Walk writes on his blog:

[I]f experimentation is gated by nine figures worth of licensing deals, we are doing a disservice to innovation … The checks being cut to ‘owners’ of training data are creating a huge barrier to entry for challengers. If Google, OpenAI, and other large tech companies can establish a high enough cost, they implicitly prevent future competition.

Now, whether there’s a barrier to entry today is debatable. Many — if not most — AI vendors have chosen to risk the wrath of IP holders, opting not to license the data on which they’re training AI models. There’s evidence that art-generating platform Midjourney, for example, is training on Disney movie stills — and Midjourney has no deal with Disney.

The tougher question to wrestle with is: Should licensing simply be the cost of doing business and experimentation in the AI space?

Walk would argue not. He advocates for a regulator-imposed “safe harbor” that’d protect any AI vendor — as well as small-time startups and researchers — from legal liability so long as they abide by certain transparency and ethical standards.

Interestingly, the U.K. recently tried to codify something along those lines, exempting the use of text and data mining for AI training from copyright considerations so long as it’s for research purposes. But those efforts ended up falling through.

Me, I’m not sure I’d go so far as Walk in his “safe harbor” proposal considering the impact AI threatens to have on an already-destabilized news industry. A recent model from The Atlantic found that if a search engine like Google were to integrate AI into search, it’d answer a user’s query 75% of the time without requiring a click-through to its website.

But perhaps there is room for carve-outs.

Publishers should be paid — and paid fairly. Is there not an outcome, though, in which they’re paid and challengers to AI incumbents — as well as academics — get access to the same data as those incumbents? I should think so. Grants are one way. Larger VC checks are another.

I can’t say I have the solution, particularly given that the courts have yet to decide whether — and to what extent — fair use shields AI vendors from copyright claims. But it’s vital we tease these things out. Otherwise, the industry could well end up in a situation where academic “brain drain” continues unabated and only a few powerful companies have access to vast pools of valuable training sets.

More TechCrunch

When VanMoof declared bankruptcy last year, it left around 5,000 customers who had pre-ordered e-bikes in the lurch. Now VanMoof is up and running under new management, and the company’s…

How VanMoof’s new owners plan to win over its old customers

Mitti Labs aims to transform rice farming in India and other South Asian markets by reducing methane emissions by 50% and water consumption by 30%.

Mitti Labs aims to make rice farming less harmful to the climate, starting in India

This is a guide on how to check whether someone compromised your online accounts.

How to tell if your online accounts have been hacked

There is a general consensus today that generative AI is going to transform business in a profound way, and companies and individuals who don’t get on board will be quickly…

The AI financial results paradox

Google’s parent company Alphabet might be on the verge of making its biggest acquisition ever. The Wall Street Journal reports that Alphabet is in advanced talks to acquire Wiz for…

Google reportedly in talks to acquire cloud security company Wiz for $23B

Featured Article

Hank Green reckons with the power — and the powerlessness — of the creator

Hank Green has had a while to think about how social media has changed us. He started making YouTube videos in 2007 with his brother, novelist John Green, at a time when the first iPhone was in development, MySpace was still relevant and Instagram didn’t exist. Seventeen years later, posting…

Hank Green reckons with the power — and the powerlessness — of the creator

Here is a timeline of Synapse’s troubles and the ongoing impact it is having on banking consumers. 

Synapse’s collapse has frozen nearly $160M from fintech users — here’s how it happened

Featured Article

Helixx wants to bring fast-food economics and Netflix pricing to EVs

When Helixx co-founder and CEO Steve Pegg looks at Daisy — the startup’s 3D printed prototype delivery van  — he sees a second chance. And he’s pulling inspiration from McDonald’s to get there.  The prototype, which made its global debut this week at the Goodwood Festival of Speed, is an…

Helixx wants to bring fast-food economics and Netflix pricing to EVs

Featured Article

India clings to cheap feature phones as brands struggle to tap new smartphone buyers

India is struggling to get new smartphone buyers, as millions of Indians don’t go for an upgrade and continue to be on feature phones.

India clings to cheap feature phones as brands struggle to tap new smartphone buyers

Roboticists at The Faboratory at Yale University have developed a way for soft robots to replicate some of the more unsettling things that animals and insects can accomplish — say,…

Meet the soft robots that can amputate limbs and fuse with other robots

Featured Article

If you’re an AT&T customer, your data has likely been stolen

This week, AT&T confirmed it will begin notifying around 110 million AT&T customers about a data breach that allowed cybercriminals to steal the phone records of “nearly all” of its customers. The stolen data contains phone numbers and AT&T records of calls and text messages during a six-month period in…

If you’re an AT&T customer, your data has likely been stolen

In the first half of 2024 alone, more than $35.5 billion was invested into AI startups globally.

Here’s the full list of 28 US AI startups that have raised $100M or more in 2024

Whistleblowers have accused OpenAI of placing illegal restrictions on how employees can communicate with government regulators, according to a letter obtained by The Washington Post. Lawyers representing anonymous whistleblowers sent…

Whistleblowers accuse OpenAI of ‘illegally restrictive’ NDAs

Business email compromise attacks are on the rise. Here’s how you can stay ahead of the hackers.

How to protect your startup from email scams

Featured Article

What exactly is an AI agent?

Regardless of how they’re defined, the agents are for helping complete tasks in an automated way with as little human interaction as possible.

What exactly is an AI agent?

Meta announced former President Donald Trump’s Facebook and Instagram accounts will no longer be subject to heightened suspension penalties, according to an updated blog post on Friday. The company says…

Meta removes special restrictions for Trump’s account ahead of 2024 elections

A Castro Valley resident was charged Thursday for allegedly slashing the tires of 17 Waymo robotaxis in San Francisco between June 24 and June 26, according to the city’s district…

Waymo cameras capture footage of person charged in alleged robotaxi tire slashings

Welcome to Startups Weekly — your weekly recap of everything you can’t miss from the world of startups. Sign up here to get it in your inbox every Friday. This…

Defending Russia’s EU neighbors

Cat-Wells said she started this platform because traditional hiring processes are exclusionary and often overlook skilled, talented disabled people.

A VC told Keely Cat-Wells to get a male, non-disabled co-founder — she balked, nabbed a $2M pre-seed round

A new study examines whether AI could be an automated helpmeet in creative tasks, with mixed results: It appeared to help less naturally creative people write more original short stories…

Experiment finds AI boosts creativity individually — but lowers it collectively

Featured Article

HeadSpin, whose founder is in prison for fraud, sold to PE firm in fire sale, sources say

In total, HeadSpin raised $117 million since its 2015 inception and was last valued at $1.1 billion in 2020.

HeadSpin, whose founder is in prison for fraud, sold to PE firm in fire sale, sources say

A bipartisan group of senators has introduced a new bill that seeks to protect artists, songwriters and journalists from having their content used to train AI models or generate AI…

New Senate bill seeks to protect artists’ and journalists’ content from AI use

When Keith Rabois announced he was leaving Founders Fund to return to Khosla Ventures in January, it came as a shock to many in the venture capital ecosystem — and…

From Ethan Choi to Spencer Peterson, venture capitalists continue to play musical chairs

Archer Aviation and Southwest Airlines are teaming up to figure out what it will take to build out a network of electric air taxis at California airports. Southwest’s customer data…

Archer’s vision of an air taxi network could benefit from Southwest customer data

If you visited the Wikipedia website on mobile this week, you might have seen a pop-up indicating that dark mode is ready for prime time.

Wikipedia’s mobile website finally gets a dark mode — here’s how to turn it on

Featured Article

What the AT&T phone records data breach means for you

The giant U.S. telco lost the information of around 110 million customers. Here’s what you need to know.

What the AT&T phone records data breach means for you

The error brings to a close SpaceX’s incredible streak of 335 flawless launches across the company’s Falcon family of rockets, which also includes the more powerful Falcon Heavy.

SpaceX Falcon 9 suffers rare failure on orbit during Starlink deployment

The AI chatbot has been trained on Amazon’s product catalog, customer reviews, community Q&As, and other public information found around the web.

Amazon AI chatbot Rufus is now live for all US customers

If X continues to violate Europe’s data protection rules, the company is on the hook for fines of up to €4,000 per day.

More bad news for Elon Musk after X user’s legal challenge to shadowban prevails

HERO Software has closed a €40 million Series B financing round, and plans to expand across Europe. 

A startup set out to fight climate change — it did it by helping plumbers