Two crawlers, one robots.txt decision, and an honest line about where our measurements stop.
Perplexity uses two separate crawlers with two separate
consequences: PerplexityBot builds the index it searches, and
Perplexity-User fetches one page while somebody is mid-question
— and blocking either one removes you from a different half of the
product. That distinction is in Perplexity’s own crawler
documentation, and it is the part of Perplexity SEO you can verify today with a
text editor and a curl command. Everything after it — which
sources get quoted, how often, in what order — is a measurement question,
and we will say plainly where this page stops being measurement.
We have not connected Perplexity. Zero calls, no data of our
own. What follows is the vendor’s documentation, checked against their doc
page, plus what we did measure on the two assistants we do run.
| User agent | What it does | Full string |
|---|---|---|
PerplexityBot | Indexes pages so Perplexity’s search can find them. Disallow this one and you are not in the index it answers from. | Mozilla/5.0 AppleWebKit/537.36 (KHTML, like Gecko; compatible; PerplexityBot/1.0; +https://perplexity.ai/perplexitybot) |
Perplexity-User | Fetches a single page while a user is asking about it. Disallow this one and the live look-up fails even though you are indexed. | Mozilla/5.0 AppleWebKit/537.36 (KHTML, like Gecko; compatible; Perplexity-User/1.0; +https://perplexity.ai/perplexity-user) |
Both rows come from Perplexity’s crawler documentation, checked on 15 August 2026 and re-checked on 5 September 2026. The complete cross-vendor version is on the AI crawler page.
If you want to be findable in Perplexity, these two rules are the floor:
User-agent: PerplexityBot
Allow: /
User-agent: Perplexity-User
Allow: /
Two things this does not do. It does not affect any other assistant — those are different tokens entirely. And it does not help if a firewall or bot-protection rule is returning 403 before robots.txt is ever consulted, which is the failure mode people miss because robots.txt looks correct while the crawler never gets a page. Send the request with the crawler’s own user agent and read the status code.
On 15 August 2026 we asked two assistants 12 questions each about one local market and kept every answer. ChatGPT attached 44 links from 25 domains; Claude attached 0 links while its transcripts show 277 pages read. The links ChatGPT showed went mostly to businesses’ own websites, with a short list of review platforms making up most of the rest.
None of that is a Perplexity number. It is reasonable to expect a citation-first product to behave more like the engine that cites than the one that does not, but expectation is not measurement, and a page that blurred the two would be doing exactly what this site was built to argue against. If you need a Perplexity figure today, the honest answer is that we cannot give you one; what we can give you is the ChatGPT run in full, with its method attached, so you can judge how far it travels.
A direct API key. The model gateway we run through does not carry a Perplexity model at all, so the Perplexity leg of our tool has never made a real call — and until it does, we count our coverage as two assistants, not three. The front page keeps that list, and the audit includes a re-run of the same questions once the missing engines are wired up. Saying “we support five platforms” before that would be the kind of claim this whole site is a complaint about.
PerplexityBot indexes pages so Perplexity's search can find them. Perplexity-User fetches a single page while a user is asking a question that needs it. They are separate tokens in robots.txt and blocking them has different consequences.
Allow both PerplexityBot and Perplexity-User in robots.txt, then send a request with each crawler's full user-agent string and confirm you get a 200. A firewall rule can refuse the crawler even when robots.txt is perfectly correct.
No. We have made zero real calls to Perplexity: the model gateway we run through does not carry it. ChatGPT and Claude are the two assistants we live-test, and a re-run is included in the audit when the others are connected.
They are different systems with different crawlers, and we have no measurement of our own that connects them. What we can say from our own run of other engines is that citation behaviour varies enough between assistants that advice does not transfer unedited.
Crawler access, because everything else is downstream of it. A page the crawler cannot fetch cannot be quoted no matter how it is written.