/
Navigation
Chronicles
Browse all articles
Explore
Semantic exploration
Research
Entity momentum
Nexus
Correlations & relationships
Story Arc
Topic evolution
Drift Map
Semantic trajectory animation
Posts
Analysis & commentary
Pulse API
Tech news intelligence API
Browse
Entities
Companies, people, products, technologies
Domains
Browse by publication source
Handles
Browse by social media handle
Detection
Concept Search
Semantic similarity search
High Impact Stories
Top coverage by position
Sentiment Analysis
Positive/negative coverage
Anomaly Detection
Unusual coverage patterns
Analysis
Rivalry Report
Compare two entities head-to-head
Semantic Pivots
Narrative discontinuities
Crisis Response
Event recovery patterns
Connected
Search: /
Command: ⌘K
Embeddings: large
TEXXR

Chronicles

The story behind the story

days · browse · Enter similar · o open

Cloudflare launches a tool that aims to block bots from scraping websites for AI training data, available free for all customers

Cloudflare, the publicly traded cloud service provider, has launched a new, free tool to prevent bots from scraping websites hosted on its platform for data to train AI models.

TechCrunch Kyle Wiggers

Context & Ripple Effects

Cloudflare had already expanded from web infrastructure into AI deployment with tools for customers to run AI models. This move addresses the other side of that ecosystem: who controls the web data those models collect.

The free blocking control became an early step in a broader policy layer, later joined by AI-bot auditing tools and a crawl-payment marketplace. That arc makes the launch consequential beyond a single bot-setting feature.

First-order effects

  • Cloudflare customers can immediately deny AI-training crawlers access to sites hosted on its network without buying a separate product.
  • AI model developers using affected crawlers can lose access to content from customers that activate the control, making site-level permission a practical constraint on collection.

Second-order effects

  • Making the control free lowers the barrier for publishers and other site operators to adopt a default-deny stance, increasing pressure on AI crawlers to identify themselves and honor site preferences.
  • The feature establishes a foundation for more granular monitoring and monetization tools, including Cloudflare's later Pay per Crawl marketplace, rather than treating crawling as a purely technical traffic-management issue.

Third-order effects

  • If widely adopted, AI training-data access shifts from an implicit open-web assumption toward a permissioned market in which infrastructure providers enforce publisher rules.
  • Control over crawler identity and access can become a strategic layer of web infrastructure, with standards and platform policies shaping which AI companies can collect data at scale.

The trend: AI-data collection is becoming an enforceable publisher-rights and infrastructure-control issue rather than an ungoverned byproduct of web crawling.

Discussion

  • @heracles@mastodon.social @heracles@mastodon.social on mastodon
    An uncommon new service from Cloudflare, to block AI Crawlers on the distribution layer of the web:  —  “We hear clearly that customers don't want AI bots visiting their websites, and especially those that do so dishonestly.  To help, we've added a brand new one-click to block al…
  • @cloudflare @cloudflare on x
    To help preserve a safe Internet for content creators, we've just launched a brand new “easy button” to block all AI bots. It's available for all customers, including those on our free tier. Read our blog post for more details: https://blog.cloudflare.com/ ...
  • @gergelyorosz Gergely Orosz on x
    I've decided to move my blog as well over to Cloudflare - I block AI crawlers in robots.txt, but clearly this is not respected. Almost all AI crawlers are a parasocial relationship, where they offer no value to any website: it's extraction in exchange for nothing (by design.)
  • @rami_bball_fan Rami Sufian on x
    @Cloudflare I'm impressed by Cloudflare's proactive move to protect content creators from AI bots! This easy button solution is a game-changer, especially with the rising concerns about AI-generated content.
  • @nfergu Neil Ferguson on x
    @Cloudflare Seems like a nice feature but I'm skeptical about this claim. How would you know if you got a false negative if you don't know the ground truth? [image]
  • @theobsguy @theobsguy on x
    Kudos to Cloudflare for this. The AI companies need to act responsibly and ethically. Hoovering up terabytes of data from the public web without the consent of content-creators does not constitute fair use. #cloudflare #llms #Ai https://blog.cloudflare.com/ ...
  • @carlhendy Carl Hendy on x
    Cloudflare has launched a new feature to block AI bots, scrapers, and crawlers with a single click, and it's free. As AI crawlers continue to swallow up web content, this tool helps protect your content from being used without consent. Many AI crawlers ignore robots txt [video]
  • @scottkrauss Scott Krauss on x
    Two things: 1. There's absolute explosion of bots crawling the web due to GenAI tools. If you want to drive revenue from these platforms you need to ensure bots can find your content 2. I think the wholesale blocking of all GenAI platforms is a mistake https://blog.cloudflare.com…
  • @mattfthompson Matt F. Thompson on x
    As robots.txt is increasingly ignored, Cloudfare launches functionality to block all AI bots. https://blog.cloudflare.com/ ...
  • @drale2k Drazen on x
    Interesting chart from Cloudflare. The top 3 AI crawlers are from Bytedance (TikTok), OpenAI and Anthropic. You can now block them with a single toggle using Cloudflare https://blog.cloudflare.com/ ... [image]
  • @rustybrick Barry Schwartz on x
    Ooo - @Cloudflare lets you block AI crawlers with one click https://blog.cloudflare.com/ ... [image]
  • @youraimarketer @youraimarketer on x
    Bytespider, Amazonbot, ClaudeBot, and GPTBot Top four AI web crawlers... Some of these AI scrapers do not respect the robots.txt file and scrape all your content. To prevent this, Cloudflare has launched a new free tool to block AI bots from scraping your website. You can [image]
  • @gergelyorosz Gergely Orosz on x
    Notice how many companies w creatives or creators as paying customers have misread what those customers want from AI. They attempt to train their own AIs on their paying customers' work, without seeking opt-in (and enraging customers.) Cloudflare, again, showing how it's done.
  • @alex @alex on x
    cloudflare doing very interesting work here https://blog.cloudflare.com/ ...
  • @iliedaboutcake @iliedaboutcake on x
    @Cloudflare Awesome, was doing this manually via WAF for ChatGPT, but this is easier :) [image]
  • @digitalbizweb @digitalbizweb on x
    Watch @Cloudflare become the most valuable company in the world. I am activating this feature right away. @thatkatieberry and friends you need to see this. https://blog.cloudflare.com/ ...
  • @jenstirrup @jenstirrup on x
    Cloudflare has launched a new feature allowing customers to block AI bots, scrapers, and crawlers. Decision-makers must now consider how this could affect the quality and breadth of AI training data, potentially impacting future #AI tools. https://blog.cloudflare.com/ ...
  • @gergelyorosz Gergely Orosz on x
    A company that “gets” its customers: Cloudflare offers functionality to block AI crawlers, which give zero value to any site (in fact, take away value: the more they crawl, the less traffic the site will later get!) They also consume resources (bandwidth, CPU, etc.) Great move:
  • r/technews r on reddit
    Cloudflare launches a tool to combat AI bots
  • r/technology r on reddit
    Cloudflare debuts one-click nuke of web-scraping AI