Cloudflare lets sites block AI training crawlers while staying searchable

djfergus · hn · 2026-09-16

Cloudflare's new blog post introduces "accountable mixed-use crawlers," addressing the long-standing dilemma where robots.txt forces an all-or-nothing choice: allow crawlers for search traffic or block them entirely.

The new approach lets site owners stay discoverable in search while explicitly disallowing the same crawlers from using content for AI training, adding accountability requirements for mixed-use bots. Discussion is ongoing on Hacker News about feasibility and crawler compliance.

Original post →

More from Infra

Infra channel →