/
Navigation
Chronicles
Browse all articles
Explore
Semantic exploration
Research
Entity momentum
Nexus
Correlations & relationships
Story Arc
Topic evolution
Drift Map
Semantic trajectory animation
Posts
Analysis & commentary
Pulse API
Tech news intelligence API
Browse
Entities
Companies, people, products, technologies
Domains
Browse by publication source
Handles
Browse by social media handle
Detection
Concept Search
Semantic similarity search
High Impact Stories
Top coverage by position
Sentiment Analysis
Positive/negative coverage
Anomaly Detection
Unusual coverage patterns
Analysis
Rivalry Report
Compare two entities head-to-head
Semantic Pivots
Narrative discontinuities
Crisis Response
Event recovery patterns
Connected
Search: /
Command: ⌘K
Embeddings: large
TEXXR

Chronicles

The story behind the story

days · browse · Enter similar · o open

Reddit CEO Steve Huffman says Microsoft, Anthropic, and Perplexity need to pay to scrape Reddit's data, and blocking them has been “a real pain in the ass”

After striking deals with Google and OpenAI, Reddit CEO Steve Huffman is calling on Microsoft and others to pay if they want to continue scraping the site's data.

The Verge Alex Heath

Context & Ripple Effects

Reddit’s position extends its earlier argument that it could not subsidize commercial use of its API and that its API was not built to support third-party apps. The company is now applying that boundary to AI and search-data collection rather than only conventional API clients.

The demand also follows reports that Reddit was restricting recent results from search engines outside Google’s indexing path, while Google and OpenAI had already secured licensing deals. That makes data access a negotiated commercial relationship, not simply an open-web default.

First-order effects

  • Microsoft, Anthropic, and Perplexity face a choice between negotiating paid access to Reddit data or losing a source they had been scraping.
  • Reddit must devote engineering and operational effort to enforcing blocks, even as it seeks to convert access into licensing revenue.

Second-order effects

  • Licensed partners gain a clearer access advantage over firms that rely on unauthorised crawling, increasing pressure on rival AI and search providers to strike comparable deals.
  • Blocking friction exposes the limits of technical access controls: platforms may need a mix of crawler restrictions, contracts, and ongoing enforcement rather than a one-time policy change.

Third-order effects

  • If more platforms follow Reddit’s approach, high-value conversational and community content could shift from broadly crawlable web material to licensed input for AI products.
  • The resulting market may favor large platforms able to negotiate and enforce access terms, while smaller publishers face a harder trade-off between search visibility and control over AI use.

The trend: AI companies’ demand for fresh web content is pushing publishers and platforms to replace passive crawling with paid, controlled data access.

Discussion

  • @alexheath Alex Heath on threads
    Reddit CEO Steve Huffman tells me it has been “a real pain in the ass” to block Microsoft's Bing, Perplexity, and Anthropic from scraping the site.  He says MSFT specifically has been using Reddit data in Bing and reselling it through an API to other search engines “without telli…
  • @jwherrman John Herrman on x
    the era of unilateral AI data scraping isn't just unethical, it's going to have enormous unintended consequences https://www.theverge.com/...
  • @carnage4life Dare Obasanjo on x
    Reddit's CEO has called out Microsoft, Anthropic and Perplexity for not paying to have the site show up in their results like Google has. Google is using their cash horde to make it harder to be disrupted by AI search and since the deal isn't exclusive, hard to claim antitrust. […
  • r/OpenAI r on reddit
    Reddit CEO says Microsoft needs to pay to search the site
  • r/technology r on reddit
    Reddit CEO says Microsoft needs to pay to search the site