Do AI crawlers obey robots.txt?

The major ones publish named user agents and state that they honour robots.txt, but the picture is not uniform:

  • Named agents can be addressed - GPTBot, ClaudeBot, PerplexityBot and Google-Extended each have their own directive.
  • Compliance is voluntary - robots.txt is a request, not an enforcement mechanism, for any crawler.
  • Blocking has a cost - a crawler you exclude cannot cite you either.
  • Few have decided - only 13% of Australian small business sites publish any AI crawler rules at all*.

Decide deliberately rather than by default: whether being quoted by an AI assistant is worth your content being read.