Cloudflare will block multipurpose AI crawlers from ad-supported pages by default
Cloudflare has announced that beginning September 15, it will block AI training and agent crawlers by default on ad-supported websites. The policy aims to protect publisher revenue models by allowing site owners to differentiate and selectively block automated scrapers while maintaining search engine indexability.
Key Takeaways
- Default blocks apply to all new domains and existing free-tier users starting September 15, 2026.
- Cloudflare defines three distinct crawler categories—Search, Agent, and Training—to allow granular per-purpose blocking.
- Research cited by Cloudflare shows imbalanced AI crawl-to-referral ratios reaching up to 50,000:1.
- Multipurpose crawlers that fail to separate search and training functions will be blocked across the network's ad-supported pages.
Why It Matters
This move effectively ends the era of 'all-or-nothing' crawler management, forcing major tech platforms to decouple search indexing from AI training. For the streaming and digital media ecosystem, this provides a technical defense against 'zero-click' traffic erosion where AI agents synthesize content without driving human visits. By leveraging its position handling roughly 20% of web traffic, Cloudflare is establishing a new baseline for digital content rights that prioritizes ad-funded business models. Industry observers should watch for how Google and Microsoft respond, specifically if they bifurcate their crawler identities to avoid losing search index visibility while maintaining AI data pipelines.
Additional Context
The shift toward aggressive crawler blocking follows significant industry data highlighting the financial toll of uncompensated scraping. Per reporting from Playwire in January 2026, research from Rutgers and Wharton found that publishers who implemented blanket blocks via robots.txt saw total traffic decline by 23%, illustrating the 'double bind' of needing search visibility while resisting AI extraction. Digital publishers including CNN and Business Insider reported 30-40% traffic drops following the widespread rollout of AI Overviews, which often satisfy user queries without a click-through. Cloudflare previously noted in July 2025 that its users were already eager for these controls, with over 1 million sites opting to block all AI bots within weeks of a manual toggle being introduced.
Simultaneously, the technical landscape has become an arms race against 'stealth' scrapers. In August 2025, Cloudflare reported that Perplexity AI was found using generic browser headers and rotating source ASNs to bypass network blocks, even when site owners explicitly disallowed access. This prompted Cloudflare to delist Perplexity as a verified bot and implement new heuristics for stealth detection. Furthermore, a Reuters report in early 2026 noted a 20% increase in scraping activity in the final quarter of 2025, particularly targeting national news outlets. In response, groups like the IAB have intensified calls for federal legislation, such as the AI Accountability for Publishers Act, to provide a legal backstop to these technical delivery controls.
Read full article at siliconrepublic.com
Enjoy our coverage?
Add StreamingMeme as a preferred source on Google to see more of our streaming news at the top of your Search results.
Add as preferred source