Skip to content
helloAISearch
How it worksFree toolsComparePricingLearnSign inStart free trial
How it worksFree toolsComparePricingLearnGlossaryAboutSign in
  1. Home/
  2. Free tools/
  3. AI robots.txt Generator

Free tool · Runs in your browser

Which AI crawlers should your robots.txt allow?

Choose allow or block for each of 18 AI crawlers, or start from a preset that allows the crawlers feeding answers and blocks the ones that only feed training. helloAISearch writes the robots.txt in your browser and re-checks it, so you see the resulting verdict for every crawler before you publish.

  • GPTChatGPTin the trial
  • CLClaudeon paid plans
  • GMGeminiin the trial
  • PXPerplexityon paid plans

Want this checked on every engine, on a schedule? Start the free 3-day trial. No credit card.

robots.txt builder
Allow answers, block training

Presets

Keeps every crawler that feeds a live answer; blocks the training-only ones.

Search indexBuilds the index an engine cites from. Block it and that engine cannot cite you.
  • OAI-SearchBotOpenAI · ChatGPT search
  • Claude-SearchBotAnthropic · Claude search
  • PerplexityBotPerplexity · Perplexity
  • DuckAssistBotDuckDuckGo · DuckAssist
User fetchFetches a page when a user asks about it. Block it and the page is invisible in live answers.
  • ChatGPT-UserOpenAI · ChatGPT (browsing)
  • Claude-UserAnthropic · Claude (browsing)
  • Perplexity-UserPerplexity · Perplexity (browsing)
Search + answersAlso runs a web search engine. Blocking it removes you from search as well as answers.
  • GooglebotGoogle · Search, AI Overviews, AI Mode
  • BingbotMicrosoft · Bing, Copilot
  • AmazonbotAmazon · Alexa
TrainingCollects training data only. Blocking it does not change today's answers.
  • GPTBotOpenAI · ChatGPT (training)
  • ClaudeBotAnthropic · Claude (training)
  • Google-ExtendedGoogle · Gemini (training)
  • Applebot-ExtendedApple · Apple Intelligence (training)
  • meta-externalagentMeta · Meta AI (training)
  • CCBotCommon Crawl · Common Crawl
  • BytespiderByteDance · Doubao
  • cohere-aiCohere · Cohere

One per line. Applied to the allowed crawlers and to the * group.

Added as a Sitemap: line at the end.

Kept unchanged above the generated block. Leave empty to get a fresh file with a * group.

How this tool works+

Your allow/block choices, disallow paths, sitemap URL and any existing robots.txt you paste, processed in your browser. Nothing is sent to a server or to an AI provider. No AI provider is called. Everything runs in your browser; nothing is sent to our server unless you ask for an email report. Rate limits apply. Details in the Privacy Policy (Free Tools sections) and the Terms of Service (section 29).

One check is a snapshot. A schedule is a measurement.

helloAISearch runs your prompts on ChatGPT, Claude, Gemini and Perplexity on a schedule, highlights every mention, and turns each gap into one move per day.

Track it daily → Start free →See pricing

Questions

About this tool

What does the “allow answers, block training” preset do?+

It allows the crawlers that feed live answers and citations (OAI-SearchBot, ChatGPT-User, Claude-User, Claude-SearchBot, PerplexityBot, Perplexity-User, Googlebot, Bingbot, Amazonbot, DuckAssistBot) and blocks the ones that only collect training data (GPTBot, ClaudeBot, Google-Extended, Applebot-Extended, meta-externalagent, CCBot, Bytespider, cohere-ai). Flip any row you disagree with.

Why does the output list the allowed crawlers explicitly?+

A crawler with no group of its own follows the * group, so listing it is not required. The generator lists it anyway so the choice is documented and survives a later edit that tightens the * group.

Will this break my existing robots.txt?+

Paste your current file into the existing field and it is kept, unchanged, above the generated block. The verification list re-parses the whole file, so you see the combined result per crawler before you publish, including any group in your file that overrides a choice.

Do crawlers obey robots.txt?+

The major vendors state that theirs do; some smaller crawlers do not. robots.txt is a request, not a wall. For a hard block use your CDN or firewall, and remember the same rule can block the crawlers that would have cited you.

Related tools

  • AI Crawler Access Checkerrobots.txt tested against 18 AI crawlers, with which blocks cost you answers
  • llms.txt Generatorspec-conformant llms.txt from your site name, summary and key pages; runs in the browser
  • GEO / AI Readiness Score0–100 AI readiness score for any public page with a fix per failed check
  • Sitemap Validatorfinds and validates sitemap.xml or a sitemap index, with every error and warning explained
  • All free tools →

Read next

  • How to get cited in AI search
  • The engines we track
  • AEO vs GEO vs SEO

Free 3-day trial · No credit card · Cancel anytime

Know what AI says about you. Every day, not once.

Three minutes from signup to reading what ChatGPT and Gemini say about your brand. Everything the trial finds is saved whether or not you subscribe.

Start free trial →See pricing

hello, AI search.

Free 3-day trial · No credit card · Cancel anytimeStart free trial →
helloAISearch
hello, AI search.

AI search visibility and AEO monitoring for solo operators and small agencies.
Operated by PulseSpark.ai LLC. Built in Pittsburgh, PA.

Product

  • How it works
  • Free tools
  • Pricing
  • AI engines
  • Compare
  • FAQ

Guides

  • Get cited in AI search
  • AI visibility tools for small agencies
  • Semrush alternatives for AI search
  • Ahrefs Brand Radar alternative
  • Track brand mentions in ChatGPT
  • Best AEO tool for SaaS startups

Learn

  • Articles
  • Glossary
  • What is AEO?
  • llms.txt

Company

  • About
  • Contact
  • Privacy
  • Terms
  • Subprocessors
© 2026 PulseSpark.ai LLC. helloAISearch is operated by PulseSpark.ai LLC.
Pittsburgh, PA