#5 · Primary category: Privacy-Preserving & Federated Data Science

ai.robots.txt

ai crawlers crawling privacy

A list of AI agents and robots to block.

Project last updated:08/26/26

GitHub Stars

4.1K

Forks

179

Contributors

75

License

MIT

Why we included this project

If your site is getting crawled by AI agents you never invited, this list gives you a quick way to push back. It tracks AI-related bots in a community-maintained registry and ships blocking rules in several formats: a standard robots.txt plus ready-made snippets for nginx, Apache, Caddy, HAProxy, and lighttpd, so you can paste in the one that matches your server instead of hand-writing rules. The accompanying bot-metrics table documents what each crawler is for, which makes the project handy as a reference when you want to know who is crawling your site. It is configuration.

Articles for this project

No articles for this project yet.

To suggest a topic or contribute an article, contact us.

Related projects in this category