Cloudflare Lets Site Owners Block AI Training Without Losing Search Indexing
Cloudflare has launched a 'Disallow AI Training' setting that allows website owners to prevent their content from being used to train AI models without affecting search engine indexing. The feature improves on the older 'Block AI Bots' control, which could inadvertently reduce a site's search visibility by treating all AI-related crawlers the same way. Cloudflare now classifies crawlers into three categories — Search, Training, and Agent — enabling more targeted access preferences. The no-training directive is published via robots.txt through Bot Preference Sync, and major crawlers from Apple, Google, and Microsoft have committed to honoring it. Training crawlers from Amazon, Anthropic, Meta, and OpenAI are expected to be blocked under this preference, while Bingbot compliance is anticipated by early 2027.
This is an AI-generated summary. ShortSingh links to the original source for the complete article.
Discussion (0)
Log in to join the discussion and vote.
Log in