AI search
How to Measure Visibility in ChatGPT and Google AI Results
How to measure AI search visibility today: Google AI Overviews and AI Mode, ChatGPT search, OpenAI crawlers, robots.txt controls and llms.txt.
SEOLOK Editorial Team · · 8 min read
Enter your domain to see what your robots.txt allows AI crawlers to do and whether your site has a valid llms.txt file. You can also generate a sample llms.txt to start from.
# Site > … ## Pages - [Site](https://example.com)
Want a free SEO report for your site?
Our team reviews your rankings, technical health and opportunities and prepares a report for you.
AI companies visit the web for different reasons and under different user-agent names. Some collect content for model training, some build an index for search answers, and some fetch a page on the spot because a user asked for it in a chat. The tool looks at two files.
1. /robots.txt: For each crawler below, it shows whether your site allows access overall:
| Token | Provider | Purpose in brief (per provider documentation) |
|---|---|---|
| GPTBot | OpenAI | Crawls content that may be used to train generative models |
| OAI-SearchBot | OpenAI | Surfaces websites in ChatGPT's search features |
| ChatGPT-User | OpenAI | Visits a page when a ChatGPT user asks for something |
| ClaudeBot | Anthropic | Collects content that may contribute to model training |
| anthropic-ai | Anthropic | A token not listed in Anthropic's current documentation but still common in robots.txt files |
| PerplexityBot | Perplexity | Surfaces websites in Perplexity search results |
| Google-Extended | Controls use of content for Gemini training and grounding; doesn't affect Google Search | |
| Applebot-Extended, CCBot, Bytespider and others | Various | Other AI-related tokens frequently seen in robots.txt |
2. /llms.txt: Does the file exist, and does it follow the basic format proposed at llmstxt.org? Under that proposal the file is Markdown, and the only required section is an H1 with the name of the site or project. It's followed by a blockquote summary, optional paragraphs of detail and H2 sections containing lists of links.
If you don't have one, the tool can generate a sample llms.txt from your site title and key pages as a starting point.
example.com).Training and search are separate decisions. OpenAI's crawler documentation explains that you can disallow GPTBot while allowing OAI-SearchBot, so content isn't used for training but can still appear in ChatGPT search. Anthropic's help article likewise defines separate tokens for ClaudeBot, Claude-User and Claude-SearchBot. Blocking everything with one line may cost you visibility you actually want.
User-initiated fetches behave differently. OpenAI notes that robots.txt rules may not apply to ChatGPT-User, and Perplexity's crawler page says Perplexity-User generally ignores robots.txt, because a person requested the fetch.
Google-Extended doesn't change your search rankings. According to Google's list of common crawlers, it's a control token only: it has no separate crawler and doesn't affect inclusion in Google Search.
Watch for accidentally blocking Googlebot. Broad rules written to keep AI bots out sometimes end up closing the site to every crawler under User-agent: *. Check the general rules in the report too.
With SEOLOK
Are you mentioned in AI answers? Measure it every week
SEOLOK asks the questions your customers might ask to ChatGPT, Gemini and Google AI Mode every week. It checks whether your name appears in the answer, whether your site is cited as a source and who is recommended, and reports the result with its likely range instead of one confident number.
According to OpenAI's documentation, GPTBot relates to content that may be used for training generative models, while OAI-SearchBot is used to surface sites in ChatGPT's search features. You can block GPTBot and still allow OAI-SearchBot. Check the provider's current documentation for the exact effect of each setting.
According to Google, no. Google-Extended manages whether content may be used for things like training Gemini models and grounding answers. It doesn't affect a site's inclusion in Google Search and isn't used as a ranking signal.
It isn't required. llms.txt is a proposed format to help language models understand a site. We can't verify which AI services read it, and we're not aware of any official statement that Google uses it for Search. It's a cheap step with an uncertain effect.
robots.txt is a request, not access control. Some providers state that robots.txt rules may not apply to fetches a user initiates directly. For a hard block you need measures at server or firewall level.
No. It only checks your access rules and your llms.txt. A feature to measure how your site appears in AI answers is being prepared and will show up in the dashboard automatically once it launches.
AI search
How to measure AI search visibility today: Google AI Overviews and AI Mode, ChatGPT search, OpenAI crawlers, robots.txt controls and llms.txt.
SEOLOK Editorial Team · · 8 min read
Rank tracking
An ordered checklist for a Google ranking drop: rule out reporting errors, technical issues, updates, competitors, seasonality and spam before you act.
SEOLOK Editorial Team · · 7 min read