User-agent: Googlebot Allow: / Disallow: /portal User-agent: Bingbot Allow: / Disallow: /portal User-agent: OAI-SearchBot Allow: / Disallow: /portal User-agent: ChatGPT-User Allow: / Disallow: /portal User-agent: PerplexityBot Allow: / Disallow: /portal User-agent: Twitterbot Allow: / Disallow: /portal User-agent: facebookexternalhit Allow: / Disallow: /portal # ── AI crawlers and answer engines ────────────────────────────────────────── # Listed explicitly rather than relying on the wildcard below. Several of these # crawlers treat an absent named rule as ambiguous, and a few operators check # for an explicit grant before including a site in generated answers. LVL1 wants # to be quotable in AI answers, so every one of them is allowed, with the portal # closed because it is private founder data. # # Two roles to keep in mind: search crawlers (OAI-SearchBot, PerplexityBot) # build the index an assistant cites, while user-triggered fetchers # (ChatGPT-User, Perplexity-User) load a page live when someone asks about it. # Both need access for LVL1 to appear in answers. # OpenAI: training, search index, and live user fetches User-agent: GPTBot Allow: / Disallow: /portal # Anthropic User-agent: ClaudeBot Allow: / Disallow: /portal User-agent: Claude-Web Allow: / Disallow: /portal User-agent: anthropic-ai Allow: / Disallow: /portal # Perplexity User-agent: Perplexity-User Allow: / Disallow: /portal # Google Gemini and AI Overviews (separate from Googlebot) User-agent: Google-Extended Allow: / Disallow: /portal # Apple Intelligence User-agent: Applebot Allow: / Disallow: /portal User-agent: Applebot-Extended Allow: / Disallow: /portal # Meta AI User-agent: meta-externalagent Allow: / Disallow: /portal User-agent: FacebookBot Allow: / Disallow: /portal # Common Crawl, which many open models train on User-agent: CCBot Allow: / Disallow: /portal # Others User-agent: Amazonbot Allow: / Disallow: /portal User-agent: cohere-ai Allow: / Disallow: /portal User-agent: YouBot Allow: / Disallow: /portal User-agent: Bytespider Allow: / Disallow: /portal User-agent: Diffbot Allow: / Disallow: /portal User-agent: Timpibot Allow: / Disallow: /portal User-agent: * Allow: / Disallow: /portal Sitemap: https://lvl1accelerator.com/sitemap.xml # Condensed, AI-readable map of the site and what each programme is. # https://llmstxt.org convention. # https://lvl1accelerator.com/llms.txt