CloudflareBryan Becker10 min readintermediate
Have it both ways: stay discoverable in search while disallowing AI training
Summary
Cloudflare adds a Disallow AI Training setting that lets sites stay indexed for search while blocking crawlers from using their content for AI model training. The feature replaces the generic Block AI Bots option with granular Search, Training, and Agent controls.
- Disallow AI Training publishes a robots.txt directive that blocks training crawlers but keeps mixed‑use bots (Applebot, Googlebot, Bingbot) allowed for search.
- The new granular controls (Search, Training, Agent) replace the old Block AI Bots setting, with automatic migration for existing customers.
- Operators must meet the "Accountable" criteria—opt‑out mechanisms, URL‑level visibility, and no impact on search—to be recognized by Cloudflare.
- Apple, Google, and Microsoft already qualify as Accountable; other major AI providers are categorized similarly for training‑only blocking.
Site owners can protect their content from being used to train AI models without losing search engine discoverability.
5/10



