Skip to content
Platform / Measure

AI crawler analytics from your server logs

Axiom GEO AI crawler analytics reads your server access logs and shows which AI bots visit your site, how often and which pages they fetch. It covers GPTBot, OAI-SearchBot, ChatGPT-User, ClaudeBot, PerplexityBot, Meta-ExternalAgent, Bytespider, CCBot, Amazonbot and more. That tells you whether AI engines can reach the pages you want them to read.

AI bots recognised, from GPTBot to Amazonbot
11+
AI bots recognised, from GPTBot to Amazonbot
see which pages each bot fetches
Per page
see which pages each bot fetches
each bot visits, from your own logs
How often
each bot visits, from your own logs

Key points

  • It reads your own server access logs, so you see real bot visits rather than estimates.
  • It recognises the main AI crawlers, including GPTBot, OAI-SearchBot, ClaudeBot and PerplexityBot.
  • You see how often each bot visits and which pages it fetches.
  • It helps check that robots.txt rules and site changes do what you meant them to.
  • A crawl is not a citation, but a page an engine never fetches is unlikely to be cited from a fresh visit.

What does AI crawler analytics answer?

It answers a question most analytics tools cannot. Are AI engines actually reading my site, and which pages? Google Analytics will not tell you, because bots do not run the tracking code. Your server logs will, because they record every request.

The trouble is that logs are long, raw text files. Few marketers want to search them by hand. This feature reads them and pulls out the AI bots for you.

How does it work?

You bring the logs and the platform does the sorting.

  1. Get your access logs. Download them from your hosting control panel or content delivery network.
  2. Load them into the platform. Upload the log files for the period you want to look at.
  3. Match the bots. Each request is checked against the user agents of known AI crawlers.
  4. Count and group. Visits are grouped by bot and by page, and counted over time.
  5. Read the results. See which bots came, how often, and which pages they fetched.

Which AI crawlers does it recognise?

It recognises the main AI bots and more besides. The table shows some of the ones you are most likely to see and who runs them.

BotRun byWhat it is broadly for
GPTBotOpenAICollecting pages that may be used to train models
OAI-SearchBotOpenAIFinding pages for ChatGPT search
ChatGPT-UserOpenAIFetching a page when a user asks ChatGPT to
ClaudeBotAnthropicCollecting pages for Claude
PerplexityBotPerplexityFinding pages for Perplexity answers
Meta-ExternalAgentMetaCollecting pages for Meta's AI work
BytespiderByteDanceCollecting pages for ByteDance
CCBotCommon CrawlBuilding an open web archive used by many AI projects
AmazonbotAmazonCollecting pages for Amazon services

Our glossary entry on AI crawlers explains the different types in more detail.

What do you see?

You see each AI bot with its number of visits over time. You can see which pages each one fetched and how often. That quickly shows whether bots reach your important pages or spend their time on old tags, filters and archive pages.

Three patterns come up often. A bot that never appears may be blocked in robots.txt, sometimes by accident. A bot that only fetches the home page may not be finding the rest of the site. And a sudden drop after a site change is often a sign that something in the new setup is turning bots away.

AI engines that search the web can only use pages they can fetch. If OAI-SearchBot or PerplexityBot never reach your service pages, those pages have little chance of being read when a buyer asks a question. Crawler data shows whether that first step is happening.

It is not the whole story. A crawl is not a citation, and a page fetched every day can still be ignored in answers. AI citation tracking shows which pages actually get cited. Crawler data sits one step before that.

It also helps you check your own rules. Many sites changed robots.txt in a hurry when AI crawlers first appeared. Our guide to AI crawlers and robots.txt explains the options. The logs show whether the rules you set are doing what you meant. If you are thinking about an llms.txt file, our llms.txt guide covers what it is and what it is not.

What should you do with what you find?

Most findings lead to one of a few simple actions.

  • If a bot you want is missing, check robots.txt and any firewall or bot protection rules.
  • If bots skip your key pages, check internal links and your XML sitemap.
  • If bots spend time on thin or duplicate pages, tidy those up or keep bots out of them.
  • If a bot you do not want is busy, decide whether to block it in robots.txt.

It helps to load a fresh set of logs after each change. That way you can see whether the fix worked rather than assume it did.

What are the limits?

A user agent is only a label, and some scrapers pretend to be well-known bots. So treat odd spikes with some caution. The data is also only as complete as the logs you load. If a content delivery network serves most requests, those logs may sit with the network rather than your server.

Who uses it?

Technical SEO leads use it after site migrations and robots.txt changes. B2B and SaaS teams use it to see whether AI bots reach their documentation and product pages. Agencies use it as part of an AI visibility audit for new clients.

Where it fits

Crawler data is the first link in the chain. Page optimisation checks that the pages bots fetch are well built. AI citation tracking shows which pages get cited in answers. AI traffic analytics shows the visitors those answers send.

Frequently asked questions

What is an AI crawler?

An AI crawler is a bot run by an AI company that fetches web pages. Some collect pages for model training, some build a search index used when the engine answers questions, and some fetch a page because a user asked the assistant to look at it. GPTBot, ClaudeBot and PerplexityBot are well-known examples.

What is GPTBot?

GPTBot is OpenAI's web crawler. OpenAI runs other bots too, including OAI-SearchBot for ChatGPT search and ChatGPT-User for pages fetched when a user asks. Each one can be allowed or blocked separately in robots.txt, which is why it helps to see which of them actually visit your site.

How do I see which AI bots visit my site?

Look in your server access logs, which record every request and the user agent that made it. Axiom GEO reads those logs for you. You load the log files and it shows which AI bots visited, how often and which pages they fetched, without you having to search the raw files.

Should I block AI crawlers?

It depends on what you want. Blocking search and user-request bots can stop your pages being fetched when an engine answers a question. Blocking training bots is a separate choice. Many businesses allow the search bots and decide on training bots case by case. Our robots.txt guide sets out the options.

Does a visit from GPTBot mean I will appear in ChatGPT?

No. A crawl means the bot fetched the page, nothing more. Whether ChatGPT names or cites you depends on many other things. The link runs the other way, though. A page no AI bot ever fetches has little chance of being read when an engine searches the web.

Where do I find my server access logs?

Most hosting control panels let you download access logs, and many content delivery networks can export them too. If you are unsure, your developer or hosting provider will know. Logs can be large, so it helps to start with a recent week or month.
Reviewed
AI crawler analytics: see GPTBot, ClaudeBot and more