What's Happening?
Cloudflare, an internet hosting giant, reports that its strategy of blocking AI crawlers by default is leading to increased licensing revenue for publishers. Cloudflare, which serves 20% of websites globally, began blocking AI crawlers last year unless
customers explicitly allow them, often due to content licensing agreements or clear value exchanges like Google search. The company initially introduced a 'pay-per-crawl' model, currently in beta testing, allowing website owners to charge AI crawlers for access. Cloudflare is now transitioning to a 'pay-per-use' model, which aims to compensate publishers when their content is directly used or surfaced in AI answers. Additionally, Cloudflare has launched an AI visibility dashboard to help publishers track how often their content appears in AI responses from platforms like ChatGPT and Claude.
Why It's Important?
This development is crucial for the U.S. publishing industry, which has grappled with the unauthorized use of its content by AI models. Cloudflare's tools create 'reliable scarcity' for publisher content, empowering them to negotiate more favorable licensing deals with AI companies. This shift could significantly alter the economic landscape for content creators, potentially leading to new revenue streams beyond traditional advertising. Neil Vogel, CEO of People Inc, noted that their ability to restrict AI scrapers using Cloudflare has made AI companies 'come to the table' for licensing. This model could particularly benefit niche publishers, including local news sites, which possess unique and valuable information that AI models seek, potentially offering them more licensing revenue than ad revenue.
What's Next?
Cloudflare plans to continue refining its 'pay-per-use' model, which is still in its early stages, to ensure publishers receive fair compensation for their content. The company is also collaborating with OpenAI on a research project to help AI search engines more effectively discover and index relevant content from participating websites, while also reducing unnecessary crawling. Cloudflare will implement a default block on mixed-purpose AI crawlers (those combining search and training) on pages with adverts starting September 15, unless tech companies separate their crawlers. This move aims to provide publishers with greater transparency and control over how their content is accessed and used by AI, fostering a more sustainable ecosystem for content creation and AI development.
Beyond the Headlines
The increasing sophistication of bot evasion techniques, where scrapers impersonate human users or Googlebot, highlights a deeper technological arms race between content providers and AI developers. Cloudflare's continuous improvement of its bot management product, Precursor, which monitors user sessions within web browsers, reflects the ongoing challenge of distinguishing legitimate AI use from unauthorized scraping. This situation also brings to light the 'leaky buckets' issue, where content syndicated to other websites can still be accessed by AI, even if the original source is protected. The broader implication is a re-evaluation of intellectual property rights in the age of AI, pushing for new frameworks that ensure creators are compensated for the value their content generates, moving beyond traditional copyright enforcement to proactive technological solutions and licensing models.











