User-agent: * Allow: / # The one endpoint under /api meant to be read by anyone but this site, and # the one llms.txt tells these same crawlers to prefer -- so it is allowed # by name before the rest of /api is closed. Allow: /api/models/roster Disallow: /api/ # /admin, /admin/faq and /admin/blog are deliberately NOT disallowed: they # carry , and a blocked URL is one Google # cannot read the noindex from — it can still list a disallowed page from # inbound links alone. Crawlable-and-noindex is what actually keeps them out. # The groups below grant nothing the wildcard above does not already grant. # They are here because silence is ambiguous: an operator reading this file # cannot tell an allow-by-default from an oversight, and several of these # crawlers are the ones that decide whether this site is quotable by an answer # engine at all. Saying it out loud is the point. /llms.txt and /llms-full.txt # are written for them specifically. # # Google-Extended is not a crawler — it is Google's opt-out for Gemini # grounding and AI Overviews, and allowing it is a deliberate choice, not an # omission. Same for Applebot-Extended. User-agent: GPTBot User-agent: OAI-SearchBot User-agent: ChatGPT-User User-agent: ClaudeBot User-agent: Claude-User User-agent: Claude-SearchBot User-agent: PerplexityBot User-agent: Perplexity-User User-agent: Google-Extended User-agent: Applebot-Extended User-agent: CCBot User-agent: meta-externalagent User-agent: Amazonbot Allow: / Allow: /api/models/roster Disallow: /api/ Sitemap: https://aiplus.pro/sitemap.xml