Reddit weighs ending $60M/year Google AI training licence as renewal negotiations stall
Reddit licenses its user-generated corpus to OpenAI and Google for AI training at roughly $60M/year from Google alone, but is threatening to walk from the Google renewal after concluding the original term was under-priced against what Gemini extracted. Q2 2026 disclosure showed data-licensing revenue of $43M, only 24% YoY growth with no new named buyers, and the stock fell more than 20%.
Where the rake sits
Reddit keeps a flat licence fee and nothing beyond it: roughly $60M/year from Google, $43M across all buyers in Q2 2026. The training happens inside OpenAI and Google, and the resulting models and the downstream value they produce, including search traffic Google no longer sends back, belong to the buyers permanently. The ceiling is set by how many frontier-model buyers exist and what they will pay for a raw input, and the only lever Reddit retains is the price of the next term, not any share of what has already been extracted.
What happened
- 2024: Reddit's IPO filing disclosed $203M in aggregate data-licensing contract value, with the Google agreement reported at approximately $60M annually.
- 2026-07-22: Reddit signalled it may walk from the Google AI-training licence at renewal, publicly framing the position as needing to 'recognise the unique value of Reddit's data' after Gemini and AI Overviews extracted value from the corpus and reduced search traffic back to Reddit.
- 2026-07-30: Reddit reported Q2 2026 total revenue of $805M, with advertising up 64% YoY to $762M and data-licensing revenue up 24% YoY to $43M (about 5% of total).
- 2026-07-30: OpenAI and Google remained the only named buyers of Reddit's licensed corpus; no major new licensing customers were disclosed.
- 2026-07-30: Reddit reported 130.3M daily active uniques and 514.6M weekly active uniques for the quarter.
- 2026-07-30: Reddit's stock fell more than 20% following the print.
- Ongoing: OpenAI CEO Sam Altman holds a significant pre-existing equity stake in Reddit, predating the licensing deal.
Who is involved
US social platform; here the party trying to reprice its AI-training licence at renewal, with the corpus as the asset.
NYSE: RDDT; Q2 2026 revenue $805M, DAUq 130.3M, WAUq 514.6M.
Alphabet's search/AI arm; here the licensee whose ~$60M/year AI-training deal with Reddit is up for renewal.
NASDAQ: GOOGL/GOOG; Alphabet 2024 revenue $350B.
US AI lab; runs a parallel Reddit licensing arrangement referenced in the same reporting.
$852B post-money valuation as of March 2026; ~11,056 employees (Aug 2026).
The reading
Same seat as the Q2 disclosure: Reddit supplies the corpus and Google does the training work in its own labs to make Gemini and Search better. Buyer-enriched.
For Google specifically the buyer's decision is Gemini/Search grounding on real human conversation, and Reddit's 514.6M weekly active uniques is one of the few live corpora that both scales and updates - which is exactly why the walk-away threat has teeth.
What has been crossing under the existing deal is a licensed content feed. What Reddit is now threatening to withhold at the perimeter is exactly that feed - a boundary drawn around the substrate rather than around any derived product.
The threat itself reads as Reddit's own diagnosis that the last licence was under-priced against the value Google extracted; per the book this stays qualitative and is not a number.
Why it matters
The renewal is the price test. If Reddit gets more for the same corpus, Google has been taking far more value out of it than $60 million a year reflects, and the 2024 contract was priced too low. If Reddit settles flat, or walks, the reading runs the other way: either the content is worth less to a model than it looked two years ago, or Google can get the same thing somewhere else. Either answer is worth a lot: it tells every owner selling an archive to a model builder what that archive is worth once the buyer has already trained on it.
The argument this deal tests: The identifier on loan
Related deals
- Reddit's capital-light data model: $1M capex against $300M+ cash flow demonstrates non-consumption and non-rivalry2026-04-30
- Chegg pivots to licensing decade of proprietary STEM academic content and expert network to AI labs2026-05-13
- Google wins $10M bankruptcy auction for Spirit Airlines' enterprise data and code to train AI models2026-08-14
- CuriosityStream Q2 2026 press release discloses $14.1M licensing revenue driven by AI training partnerships2026-08-12
Sources
- cnbc.com — primary
- cryptobriefing.com — corroborating
- businesswire.com — party background
- en.wikipedia.org — party background
- en.wikipedia.org — party background
- sacra.com — party background
Announced 2026-07-22 (reported) · Added to the register 2026-09-30