Skip to content
Ranqo
Docs
DashboardGet started
GuidesMethodologyIntegrations
Video tutorials

Metrics

  • How Ranqo measures
  • Visibility
  • Share of Voice
  • Average position
  • Sentiment
  • Citations
  • Source and article types
  • Difficulty and volume

Trends

  • One point per run
  • Carried values
  • Time zones

Measuring impact

  • Measurement windows
  • Noise bands
  • Confirmed wins

Data

  • Competitor identity
  • Why numbers change
  • AI crawlers

8 sections

Data

Crawlers

View as MarkdownOpen this page as .md
ChatGPTOpen in ChatGPTAsk about this pageClaudeOpen in ClaudeAsk about this page

RanqoBot, Ranqo's own crawler, and how to allow or block it, plus the AI crawlers Site Access tests and the bots Site Tracking recognises.

Crawlers come up in Ranqo in two ways. Ranqo has its own crawler, RanqoBot, which fetches pages for some features. And two features work with the AI crawlers that visit your site: Site Access tests whether they can fetch it, and Site Tracking counts their real visits. Each of those two features keeps its own list, and this page says which is which.

RanqoBot#

RanqoBot is the crawler Ranqo uses whenever it follows links or fetches in volume. Its User-Agent contains the token RanqoBot and a link to ranqo.ai/bot, the page for site owners that describes exactly what it does. If you found RanqoBot in your access logs, that page is the full answer.

Each of its requests traces to someone using Ranqo:

  • Page discovery. When a Ranqo customer audits a site, RanqoBot follows internal links on that one site to find pages worth auditing.
  • Outbound link checks. When an audited page links to another site, RanqoBot checks the link still works, for up to 50 links per audit.
  • Citation sources. When an AI engine cites a page, RanqoBot reads its title, or follows the engine's redirect link to the real address, so the publisher can be named.
  • Site Access. When Site Access checks a Ranqo customer's site, RanqoBot reads its robots.txt, sitemap and llms.txt.
  • robots.txt. Before any of the above, and before Site Access or the AI Crawler Inspector fetches a page, it reads the site's robots.txt.

Allow or block RanqoBot#

RanqoBot obeys robots.txt, including Allow rules, wildcards and patterns anchored to the end of a path. A group that names RanqoBot takes precedence over your * group. To block it entirely:

robots.txt
User-agent: RanqoBot
Disallow: /

Ranqo keeps a copy of each site's robots.txt for 12 hours, so a change can take that long to apply. Refusing the User-Agent at your server or CDN takes effect at once. The RanqoBot page has more examples and the full detail.

Other ways Ranqo fetches a page#

Not every request from Ranqo is RanqoBot, and a robots.txt rule for RanqoBot does not stop the others:

  • One page on request. When someone pastes a URL into a page audit or most free tools, Ranqo fetches that single page with an ordinary browser User-Agent, follows no links, and queues nothing.
  • The free llms.txt generator reads your homepage and sitemap under its own User-Agent, RanqoLlmsTxtBot.
  • Each crawler's own User-Agent. Site Access and the free AI Crawler Inspector send GPTBot's, ClaudeBot's and the other crawlers' own User-Agents, to show how your server answers each one. These requests come from Ranqo, not from the AI vendors, and neither tool runs if your robots.txt disallows RanqoBot. xAI does not publish its crawler's User-Agent or robots.txt token, so the xAI check uses ones Ranqo assumes (xAI), and its result says how your server treats that string rather than xAI's real crawler.

The RanqoBot page lists every User-Agent Ranqo sends.

Ranqo's requests in Site Tracking#

If your site runs Site Tracking, your install reports Ranqo's own requests like any other visit. Site Access checks your site on its own after your tracking runs, at most once every 7 days, so these requests arrive whether or not you open it.

  • Site Access and the AI Crawler Inspector send each crawler's User-Agent from Ranqo's servers. On the Traffic page, those requests appear as visits from the crawlers they name, except DuckAssistBot and xAI, which Site Tracking's catalog does not list, so they count as people. Where Ranqo has a crawler's published IP ranges, the visit is marked unverified, because Ranqo's servers are not in them. The browser request each check compares against counts as a person.
  • RanqoBot and RanqoLlmsTxtBot match no bot in the catalog either, so their requests, including RanqoBot's reads of robots.txt, sitemaps and llms.txt, count as people. So do the single-page fetches Ranqo makes with a browser User-Agent.

If you block RanqoBot on your own site#

If you are a Ranqo customer and your robots.txt disallows RanqoBot:

  • Site Access reads your robots.txt and then stops before fetching any page, says so, and shows the verdict Unknown.
  • Page discovery cannot follow links on your site to find pages to audit.
  • An audit of a single URL you enter still runs, because that one fetch is not RanqoBot.

Which crawler list belongs to which feature#

FeatureWhat it does with crawlersThe list it uses
Site Access, the free AI Crawler Inspector and the free Robots.txt GeneratorSite Access and the Inspector fetch your site as each crawler and read what your robots.txt says to each; the Generator writes robots.txt rules for eachThe AI crawler catalog below
Site Tracking, on the Traffic pageClassifies the real visits your server reports, by User-AgentSite Tracking's own catalog

The two lists overlap but are not the same, because the jobs differ. Site Access needs the crawlers that decide whether AI can read you. Site Tracking also has to recognise traffic that is not AI at all, so it is not counted as people. The lists are kept separately, so a crawler Site Access tests is not always one Site Tracking recognises: DuckAssistBot and xAI are in the first list only.

The crawlers Site Access tests#

A Site Access check fetches your homepage, and a few pages picked from your sitemap (up to 4 pages in all), as each of the 16 crawlers marked fetchable below, and as a regular browser to compare against. Some entries are permissions only: robots.txt tokens that a vendor honours but no crawler sends as a User-Agent. For those, Site Access reports what your robots.txt says, because that is the whole answer.

The User-Agent Site Access sends for each crawler is an approximation modelled on that crawler's own, and contains its token.

AI crawlers
CrawlerVendorRolerobots.txt tokenFetchable
GPTBotOpenAITrainingGPTBotYes
ChatGPT-UserOpenAIAssistant (live fetch)ChatGPT-UserYes
OAI-SearchBotOpenAIAI search indexOAI-SearchBotYes
ClaudeBotAnthropicTrainingClaudeBotYes
anthropic-aiAnthropicTraininganthropic-aiYes
Claude-UserAnthropicAssistant (live fetch)Claude-UserYes
Google-ExtendedGoogleTrainingGoogle-Extendedrobots.txt token only
GooglebotGoogleSearch engineGooglebotYes
PerplexityBotPerplexityAI search indexPerplexityBotYes
Perplexity-UserPerplexityAssistant (live fetch)Perplexity-UserYes
Applebot-ExtendedAppleTrainingApplebot-Extendedrobots.txt token only
ApplebotAppleSearch engineApplebotYes
BytespiderByteDanceTrainingBytespiderYes
CCBotCommon CrawlTrainingCCBotYes
BingbotMicrosoftSearch engineBingbotYes
DuckAssistBotDuckDuckGoAssistant (live fetch)DuckAssistBotYes
Meta-ExternalAgentMetaTrainingMeta-ExternalAgentYes
xAIxAITrainingxAIYes

The Crawlers tab in Site Access groups them by what they are for:

  • Live answers: assistants that fetch a page while answering someone, and AI search indexes. A block here can take you out of AI answers.
  • Search: traditional search engine crawlers.
  • Training: crawlers that collect data to train future models. Blocking one is reported as a licensing choice, never as a blocker: AI answers cite pages through the live-answer and search crawlers.

See Site Access for what each result means.

The bots Site Tracking recognises#

Site Tracking checks the User-Agent of every visit your server reports against its own catalog, and puts each recognised bot in one of these groups on the Traffic page:

  • Training: crawlers collecting data for AI model training.
  • Indexing: AI search indexes that decide what an assistant can cite.
  • Agentic: fetches by an AI assistant acting on someone's request.
  • Search Engine: traditional search engine crawlers.
  • Social Preview: link-preview generators from chat and social apps, and automated browsers such as performance audits and headless test browsers.
Bots Site Tracking recognises
BotVendorCategoryIP verifiable
GPTBotOpenAITrainingYes
OAI-SearchBotOpenAIIndexingYes
ChatGPT-UserOpenAIAgenticYes
ClaudeBotAnthropicTrainingYes
Claude-UserAnthropicAgenticYes
Claude-SearchBotAnthropicIndexingYes
PerplexityBotPerplexityIndexingYes
Perplexity-UserPerplexityAgenticYes
GooglebotGoogleSearch EngineYes
Google-ExtendedGoogleTrainingNot applicable: robots.txt token only
GoogleOtherGoogleIndexingYes
bingbotMicrosoftSearch EngineYes
BingPreviewMicrosoftSocial PreviewYes
Meta-ExternalAgentMetaTrainingNo
Meta-ExternalFetcherMetaAgenticNo
facebookexternalhitMetaSocial PreviewNo
ApplebotAppleSearch EngineNo
Applebot-ExtendedAppleTrainingNot applicable: robots.txt token only
CCBotCommon CrawlTrainingNo
BytespiderByteDanceTrainingNo
cohere-aiCohereTrainingNo
DiffbotDiffbotIndexingNo
AmazonbotAmazonSearch EngineNo
YouBotYou.comIndexingNo
SlackbotSlackSocial PreviewNo
DiscordbotDiscordSocial PreviewNo
TwitterbotXSocial PreviewNo
LinkedInBotLinkedInSocial PreviewNo
WhatsAppMetaSocial PreviewNo
TelegramBotTelegramSocial PreviewNo
Chrome-LighthouseGoogleSocial PreviewNo
HeadlessChromeGoogleSocial PreviewNo
PlaywrightMicrosoftSocial PreviewNo
PuppeteerGoogleSocial PreviewNo
CypressCypressSocial PreviewNo

Rows marked robots.txt token only are permissions, not crawlers. No crawler sends them as a User-Agent (see the Site Access table above), so a visit that carries one is someone else using the name.

A visit whose User-Agent matches no bot is counted as a person: Human (AI Referral) when it arrived from an AI assistant's site, and Human (Direct) otherwise. That includes crawlers the catalog does not list, such as RanqoBot, DuckAssistBot, xAI and Googlebot-Image.

When a bot's vendor publishes the IP addresses its crawlers use, Ranqo checks each visit against them. On the Traffic page, the Verified column of Agents detected shows the share of a bot's checked visits that came from those addresses. A visit that claims to be a bot from anywhere else may be someone else using its name. Automated browsers cannot be verified this way, because anyone can run them.

See Site Tracking for installing it and how visits are counted.

Previous page: Why numbers changeNext page: Integrations
On this page
Ranqo
Docs
Dashboardranqo.aiPrivacyTerms
Be the Source AI Cites.
Ranqo
Docs
GuidesMethodologyIntegrations

Metrics

  • How Ranqo measures
  • Visibility
  • Share of Voice
  • Average position
  • Sentiment
  • Citations
  • Source and article types
  • Difficulty and volume

Trends

  • One point per run
  • Carried values
  • Time zones

Measuring impact

  • Measurement windows
  • Noise bands
  • Confirmed wins

Data

  • Competitor identity
  • Why numbers change
  • AI crawlers
Get started