AI-Economy

Cloudflare separates AI training from search engine access

3 min read

TL;DR Too Long; Didn’t read

Since September 15, 2026, Cloudflare has separated the Google visibility of a website from the permission for AI training. The new switch 'Disallow AI Training' requires the 'Accountable' status, which currently seven providers hold. New, ad-funded sites will automatically block AI training in the future. Only 17 percent of customers are currently actively using such blocks.

An archway with the Cloudflare logo as a sticker above its peak splits into two paths behind it: one leads to a large magnifying glass, the other ends at a red barrier in front of a stylized robot head. Image generated with GPT Image 2

Key takeaways

  • The new switch 'Disallow AI Training' completely replaces the previous blanket bot blocker at Cloudflare.
  • Apple, Google, and Microsoft receive the 'Accountable' status for their combined search-and-training crawlers.
  • Amazon, Anthropic, Meta, and OpenAI are also considered 'Accountable' for their separate training crawlers.
  • Less than one percent of Cloudflare customers currently block search engine crawlers, 17 percent block training.
  • Mixed crawlers that combine search and training account for 36.6 percent of bot traffic, according to Cloudflare.
  • Cloudflare plans to integrate the opt-out of AI summaries directly into its own dashboard by early 2027.

Cloudflare introduced a new setting called “Disallow AI Training” on September 15, 2026: website operators can refuse AI training without disappearing from Google search. The prerequisite is the new “Accountable” status, which is currently held by seven major tech companies. Newly created, ad-funded sites will automatically block AI training in the future.

New switch replaces blanket bot blocking

Until now, Cloudflare customers only had a rough choice: block or allow AI bots. The new switch has four levels – Allow, Disallow AI Training, Block on Ad Pages, and Completely Block – and can be set individually for each domain, as Cloudflare explains in its blog post. In July, the company had already announced that it would block training and agent bots by default on ad-supported pages starting September 15. Newly added, ad-funded domains will now automatically receive the “Disallow AI Training” setting as a default, while ad-free pages will continue to allow all bot types. Existing customer settings were automatically converted by Cloudflare to the new category by the deadline, ensuring that search engines like Google and Apple remained crawlable while pure training accesses were eliminated, reports Search Engine Journal. CEO Matthew Prince justifies the move by stating that Cloudflare wants to maintain the openness of search while giving rights holders real control over the use of their content.

Accountable status applies to seven major tech companies

Choosing the “Disallow AI Training” setting does not automatically result in a loss of visibility in search – provided the respective provider holds the new Accountable status. To obtain it, a company must meet four conditions according to Cloudflare: an opt-out option from AI training via robots.txt or a comparable standard, an opt-out option from AI-generated summaries, initially directly with the respective provider and from early 2027 bundled in the Cloudflare dashboard, insight at the URL level into training usage, and assurance that an opt-out from training does not worsen search ranking. Apple, Google, and Microsoft meet these criteria for their combined search-and-training crawlers like Googlebot and Applebot. Amazon, Anthropic, Meta, and OpenAI are also considered Accountable for their separately operated training crawlers. Therefore, anyone who does not block any of these seven providers can refuse AI training without disappearing from Google or Bing results – for all other unlisted mixed crawlers, however, the strictest chosen rule applies.

Numbers show cautious use of the new controls

Despite the new options, only a few website operators are currently using granular bot controls: According to Cloudflare’s own, independently unverified network data, less than one percent of all customers block search engine crawlers, while 17 percent restrict AI training. Mixed crawlers, which bundle search and training in a bot identifier, account for 36.6 percent of verified crawler traffic – exactly the group for which the new Accountable status is intended. Cloudflare justifies the effort with user behavior: More than 40 percent of people who read an AI summary abandon their search afterward, while visitors coming through AI search services convert three to five times more often than traditional search visitors. For operators of ad-funded sites, this creates a conflict of interest between visibility in new AI search paths and protecting their own content from unpaid training.

It remains to be seen whether the currently low usage rate of 17 percent will increase once more operators learn about the new fine-tuning – or whether the complexity of four bot categories and four Accountable criteria will deter many. A real litmus test will come only in early 2027 when Cloudflare plans to integrate the opt-out from AI summaries into its own dashboard: only then could the entire range of AI usage – training, search, and automatically generated responses – actually be controlled in one place.

Frequently asked questions

What does the new setting cost for Cloudflare customers?

Cloudflare does not specify a separate price for 'Disallow AI Training'; the function is based on the bot categories introduced in July 2026, which are available for all tariff levels, including the free plan.

Does the new regulation also apply to websites in Germany and the EU?

Yes, the settings are available worldwide through the Cloudflare dashboard and are not limited to individual regions; Cloudflare does not mention an EU-specific variant.

What happens to crawlers from companies without Accountable status?

If an operator selects 'Disallow AI Training', Cloudflare completely blocks mixed crawlers without this status – they also lose access for search because the system treats them according to the strictest applicable rule.

How does this differ from the rule announced in July?

In July, Cloudflare had only announced a blanket blockade of training and agent bots on advertising pages by September 15; the now-launched fine-tuning allows operators for the first time to specifically refuse training without losing search visibility.

Can operators change the new default settings?

Yes, new and existing customers can individually adjust any of the four levels at any time in the security area of the dashboard.

Sources (3)
  1. Cloudflare Blog: Have it both ways: stay discoverable in search while disallowing AI training
  2. Cloudflare Press Release: Cloudflare Helps End the Search-or-AI-Training Tradeoff
  3. Search Engine Journal: Cloudflare Lets Sites Disallow AI Training Without Blocking Googlebot

Your AI update for the work week

Once a week, the most important AI news – plus one practical tip to try right away. No spam, unsubscribe anytime.

← Back to the blog