AI crawler list: 81 crawlers, who runs each, and how to allow or block it

Every AI crawler Citable recognises: who runs it, when it reads, its name in robots.txt, whether its identity can be checked, and how to allow or block it on 12 platforms.

This response

You are ClaudeBot (Anthropic). Citable recognised you as an AI agent, so this is https://getcitable.in/crawlers with the UI removed — the content only. A browser asking for the same address gets the full designed page.

The 81 crawlers

This is the list Citable's own edges match on. Googlebot and Bingbot are not on it: search crawlers are never treated as AI crawlers, and always get the real page.

By platform

Two things a robots.txt rule does not do

A group that names a crawler replaces the * group for that crawler; it does not add to it. Paths the * group closes are open to a crawler with its own group unless they are repeated there.

robots.txt is a request. A crawler that honours it stays out; nothing in the file stops one that does not. What is enforced is what the server or the edge in front of it refuses.

About Citable

Citable strips the UI from a website for AI agents and serves them only the content.

When an AI agent — ChatGPT, Claude, Perplexity or any of the 81 that Citable recognises — opens a page, Citable removes the theme, scripts, navigation and layout and serves only the content: the same facts, at the same URL, on the site's own domain. People and Google still get the full designed page.

How it works: https://getcitable.in/how-it-works · Summary for AI agents: https://getcitable.in/llms.txt


Source: https://getcitable.in/crawlers · Citable · Run the free audit · llms.txt · getcitable@gmail.com