Skip to content
siterank.info

Knowledge hub

AI crawlers

The bots behind AI search, training corpora and user-triggered fetches, and how robots.txt does and does not govern them. Under RFC 9309 those rules are requests honoured voluntarily, not access control.

Guides in this topic

Filtered to this topic. Use the row above to widen or change it.

PerplexityBot and Perplexity-User Explained

Perplexity runs two agents, and states that one of them generally ignores robots.txt. Here is what each does and how to configure for it.

Published September 9, 2026 Reviewed September 9, 2026 Sources checked September 9, 2026 AI crawlers

Should You Block AI Training Crawlers?

A licensing decision, not an SEO one. What training crawlers are, what blocking actually changes, and a framework for deciding by site type.

Published September 9, 2026 Reviewed September 9, 2026 Sources checked September 9, 2026 AI crawlers

What Is llms.txt? What It Can and Cannot Do

llms.txt is a community proposal for a Markdown site map aimed at LLMs. Here is the format, its real status, and what publishing one does not do.

Published September 9, 2026 Reviewed September 9, 2026 Sources checked September 9, 2026 AI crawlers