Skip to main content
Tools / 01 · Site audit Free · tests 17 bots · result in-page
FreeSite audit

Can AI bots find your content?

Check access policies for 17 crawlers, including OpenAI and Claude search, browsing, and training agents.

One free audit per email. Verify with a 6-digit code to view it.

01 / What it does

Ranking on Google says nothing about AI access.

AI search engines like ChatGPT, Perplexity, and Gemini use their own web crawlers to index content. Each bot has a unique user-agent, and many websites accidentally block them through robots.txt rules written for traditional search.

An AI crawl check audits your site against 17 crawler user-agents, separates retrieval access from training policy, and flags missing structured data and llms.txt files that help AI models understand your content. Without this visibility, your site may be invisible to AI search results even if it ranks well on Google.

Search crawlers and training crawlers do different jobs

Crawler typeExamplesWhat it is forWhat the vendor says about blocking it
TrainingGPTBot, ClaudeBotCollect content that may be used to train AI modelsAnthropic: restricting ClaudeBot signals that your future content should be excluded from its training data.
SearchOAI-SearchBot, Claude-SearchBot, PerplexityBotSurface and link pages in the assistant's search resultsOpenAI: sites that disallow OAI-SearchBot will not be shown in ChatGPT search answers, though they can still appear as navigational links. Anthropic: blocking Claude-SearchBot may reduce your visibility in user search results. Perplexity: PerplexityBot is not used to train AI foundation models; its page does not say what blocking does.
User-triggeredChatGPT-User, Claude-UserFetch a page when a person asks the assistant toOpenAI: robots.txt rules may not apply to ChatGPT-User, because a person starts the fetch. Anthropic: disabling Claude-User may reduce your visibility for user-directed web search.

To stay in ChatGPT search answers, make sure your robots.txt does not disallow OAI-SearchBot. An explicit rule looks like this:

User-agent: OAI-SearchBot
Allow: /

Sources: OpenAI crawler docs, Anthropic's crawler article and Perplexity's bot docs, checked 26 Sep 2026.

02 / What we check

What we check.

4 checks, each with a plain-English reason it matters. Every one maps to a line in the report.

01AI Crawl Checker

17 Bot User-Agents

HEAD requests as Googlebot, OAI-SearchBot, Claude-SearchBot, PerplexityBot, GPTBot, and more.

02AI Crawl Checker

robots.txt Analysis

Parse allow/disallow rules per bot, sitemap directive, and crawl-delay.

03AI Crawl Checker

Structured Data

JSON-LD extraction: Article, FAQPage, Organization, Product, BreadcrumbList.

04AI Crawl Checker

llms.txt Detection

Check for the emerging standard that helps AI understand your site.

03 / Scoring

How We Score (And Why It's Different)

100points

Most crawl checkers stop at robots.txt. We score what actually determines whether AI systems cite you: structured data quality, llms.txt presence, and content accessibility. This methodology comes from our own work optimizing for AI search.

01
Bot Access & robots.txt

Do AI retrieval bots (OAI-SearchBot, Claude-SearchBot, PerplexityBot) have access? Is there a deliberate policy for training crawlers? A smart robots.txt isn't just allow/block. It's a strategy.

40pts
02
Structured Data

JSON-LD is how machines understand your content. We check for page-type-appropriate schema: Organization for homepages, Article for posts, BreadcrumbList for navigation. FAQPage also scores, though Google stopped showing FAQ rich results on May 7, 2026.

25pts
03
llms.txt

A proposed file that summarizes your site for language models. We score structure, entity definitions, URLs, and use policy. It is optional: Google says Google Search ignores it, and no major AI engine documents using it to choose citations.

20pts
04
Content Quality

Can AI crawlers actually read your content without executing JavaScript? We check server-side rendering, title/description length, and framework detection.

15pts
Radar

Want the full picture across four AI engines?

Radar orchestrates 13 AI visibility audits in staged batches: crawl access, llms.txt, structured data, brand citations, and more. It then surfaces cross-tool conflicts and ships a prioritized action plan for your domain.

Sweep · 000° · 13 checks · one pass
Start now

Run it on your site. See what AI can read.

Free, no account, result on this page. When you want the whole picture, Radar runs all 13 audits in one pass.