AI news story
Cloudflare replaces its blanket AI bot block with granular controls for search, training, and agent crawlers
Cloudflare is giving all customers granular AI bot controls. Site owners can now manage Search, Training, and Agent bots separ…
Editor's take
Cloudflare has shifted from a broad AI bot blocking strategy to a more nuanced system, allowing website owners to differentiate between search engine crawlers, AI training bots, and other agent-based traffic. This move acknowledges the growing distinction between beneficial AI-driven discovery and potentially resource-intensive or data-scraping AI applications. For businesses relying on web traffic analytics and resource management, this offers greater precision in managing how AI systems interact with their online presence, directly impacting their infrastructure costs and data integrity.
The default blocking of Training and Agent bots from September 15, 2026, signals a clear industry trend towards stricter control over AI data ingestion. This will likely force AI developers, particularly those training large models like OpenAI's GPT series or Google's Gemini, to actively seek opt-in mechanisms or negotiate access. The next critical development to observe will be the adoption rate of these granular controls by website owners and the subsequent impact on the availability and cost of training data for AI companies.