Reddit blocks the Internet Archive from crawling its data - heres why

General News

Summary

Reddit is blocking the Internet Archive’s Wayback Machine from crawling most of its pages. The company says it wants to stop AI firms from using archived Reddit content to scrape user data and train models. The new limits allow crawling only the homepage and block access to user profiles, comments, and post detail pages. The move adds to broader tension between content owners and AI developers over training data and licensing rights.

Classifications

industries
HealthTech
applications
Accounting and Taxes

AskAI Classifications

Labels
SaaS Consumer Software Enterprise Software

Linked Companies

Google LLC
$100M to $250M
OpenAI
$25M to $50M
Anthropic
$10M to $25M