How to Rank in Perplexity: Optimizing for an Index Nobody Else Crawls
Your Bing Webmaster Tools account is verified, your Google Search Console sits clean, and you still cannot work out how to rank in Perplexity. Neither of those two accounts reports on Perplexity at all, because Perplexity does not retrieve through either index. Citant.ai starts every attempt to optimize for Perplexity from Perplexity’s own crawler documentation, which names PerplexityBot and Perplexity-User as two separate agents with two different robots.txt behaviors.
Perplexity optimization is the practice of making a page crawlable by PerplexityBot, present in Perplexity’s own search index, and quotable enough that Perplexity names the brand inside its numbered source list. Perplexity operates its own crawler, PerplexityBot, and its own search index. Citant treats Perplexity indexation as its own workstream inside the 4-Layer Citation Framework, checked at the Retrieval layer before any page is rewritten. Citant.ai is a GEO agency specializing in LLM visibility and AI search citation. Perplexity AI is one of the six platforms tracked in that work, and it is the only one of the six sitting on an index that no other tracked platform shares.
To rank in Perplexity, a page must be crawlable by PerplexityBot, present in Perplexity’s own search index, and written so a single passage answers one question completely.
Perplexity publishes no submission route and no inclusion request in its crawler documentation. Bing Webmaster Tools work and Google Search Console work reach neither the crawler nor the index.
On this page
- How Perplexity work fits inside Citant’s GEO practice
- PerplexityBot and Perplexity-User are two different agents
- Why Bing and Google work does not reach Perplexity
- Where the crawler token detail lives
- What a Perplexity citation actually looks like
- The drafting standard that follows from numbered citations
- Why a Google ranking is not a Perplexity citation
- Where Perplexity’s citations come from
- What the evidence says about page structure on Perplexity
- The ghost citation problem on Perplexity
- Sponsored placement is not a citation
- The Perplexity optimization checklist
- How Citant.ai measures whether Perplexity names a brand
- What nobody can promise on Perplexity
- Key takeaways
- Frequently asked questions
- Related guides
How Perplexity work fits inside Citant’s GEO practice
Perplexity work sits inside a GEO engagement as a platform workstream, never as a standalone product. Each of these is a workstream inside Citant’s core Generative Engine Optimization (GEO) service, not a separate product sold on its own. Content can pass retrieval and re-ranking and still fail to get the brand named in the final answer, a failure Citant calls a ghost citation. For a B2B SaaS marketing team that needs to know whether Perplexity names the brand at all before deciding what to change, AI visibility services report Share of Model on each tracked platform separately. For teams that want Perplexity handled alongside the other five platforms rather than as a side project, our GEO service runs the 4-Layer Citation Framework across all six.
PerplexityBot and Perplexity-User are two different agents
The two Perplexity user agents do two different jobs, and a robots.txt rule that treats them as one produces the wrong outcome twice. Perplexity documents PerplexityBot as the agent behind its search results, stating that it “is designed to surface and link websites in search results on Perplexity” and that it “is not used to crawl content for AI foundation models”. Perplexity documents Perplexity-User as a reader-triggered fetcher instead, stating that it “supports user actions within Perplexity” and that “since a user requested the fetch, this fetcher generally ignores robots.txt rules”. Source: Perplexity crawler documentation, fetched 23 August 2026. Last verified: August 2026.
| Perplexity user agent | What it does | Respects robots.txt | What blocking it costs a publisher |
|---|---|---|---|
| PerplexityBot | Surfaces and links websites in search results on Perplexity, and is not used to crawl content for AI foundation models | Yes. Perplexity recommends allowing it in robots.txt so a site appears in search results | Index absence, which removes the site from the pool Perplexity answers can cite at all |
| Perplexity-User | Visits a web page when a reader asks a question, to help answer it and link the page in the response | No. Perplexity states this fetcher generally ignores robots.txt rules because a user requested the fetch | Little, since the fetch follows a reader’s question rather than building the index |
Why Bing and Google work does not reach Perplexity
Perplexity indexation is a separate technical job from Bing indexation and Google indexation, and no amount of work in the other two consoles substitutes for it. Perplexity publishes IP ranges for both of its agents, in a published PerplexityBot IP range file and a published Perplexity-User IP range file, so a publisher can confirm a real visit in server logs rather than trusting a user agent string. A firewall or CDN rule that blocks unrecognized bots stops PerplexityBot with no robots.txt evidence at all. Teams that want crawler access checked against a named list before rewriting anything can run the AI Crawler Access Report, a free tool from Citant.ai.
The free scan checks whether AI systems can reach the site. The $440 AI Search Audit measures whether they name the brand.
Where the crawler token detail lives
Token-level robots.txt syntax for every AI crawler belongs in one place on this site rather than being restated per platform. This page states only the PerplexityBot and Perplexity-User distinction, because that distinction is what changes a Perplexity decision. For the directive syntax, the full token list, and the vendor documentation dates behind each entry, the full AI crawler reference table carries the complete set across every named AI crawler.
What a Perplexity citation actually looks like
A Perplexity citation is a numbered inline marker attached to a visible source list, not a link buried under a summary. Perplexity states that “each answer includes numbered citations linking to the original sources, allowing you to easily verify the information or explore further”, per the Perplexity Help Center, fetched 23 August 2026. Last verified: August 2026. That format sets the unit of work: the thing being cited is a passage a reader can check against the claim, so a page is scored by its passages rather than by its total length.
The drafting standard that follows from numbered citations
Citant’s drafting standard is a self-contained block of 50 to 150 words that answers exactly one question and reads correctly with nothing around it.
That range is a house drafting standard built on how retrieval-augmented systems split documents into retrieval units, and it is never presented as a measured citation multiplier. Every heading opens with a 40 to 75 word direct answer to the question the heading asks, before any setup. A passage lifted into a numbered source list carries no antecedent with it, which is why no factual sentence on a Perplexity-targeted page should open with an unanchored pronoun.
Why a Google ranking is not a Perplexity citation
Perplexity overlaps Google’s top 10 more than the other tracked assistants do, and still nowhere near enough to treat one as a proxy for the other. In a University of Toronto study of 1,000 consumer ranking queries, the mean overlap between the domains Perplexity Sonar Pro cited and Google’s top 10 results was 15.2%, with a median of 14.3% (Chen et al., EDBT/ICDT 2026 Workshops), published as Navigating the Shift. Two bounds travel with that figure: the corpus is consumer ranking queries across ten retail and travel categories rather than B2B SaaS, and the paper reports no explicit calendar collection window. Read against a Perplexity plan, it says the same thing the crawler documentation says: a separate index needs separate work.
Where Perplexity’s citations come from
Perplexity spreads its citations across earned, brand-owned and social sources differently from every other assistant measured alongside it. In a University of Toronto preprint from the same research team, Perplexity allocated 73.4% of its citations for niche brands to earned sources, 9.1% to brand-owned domains and 17.5% to social sources (Chen et al., September 2025), published as Generative Engine Optimization.
Three bounds travel with that figure: it is a preprint with no peer review, the corpus is consumer brand queries rather than B2B SaaS, and the classification is LLM-assisted. Social sources appear frequently in Perplexity answers, which is an observed citation pattern in output analysis rather than a stated part of Perplexity’s retrieval architecture.
What the evidence says about page structure on Perplexity
Perplexity is one of only three engines in the measured literature whose citations have been audited against a page-level scoring framework on B2B SaaS pages specifically. A GEO-16 study collected 1,702 citations from 70 product-intent prompts and audited 1,100 unique URLs across Brave Summary, Google AI Overviews and Perplexity, finding that the pillars covering metadata and freshness, semantic HTML and structured data showed the strongest associations with citation (Kumar and Palkhouski, September 2025). Three cautions travel with that finding: the paper is a preprint with no peer review, it reports pillars that align with citation rather than cause it, and it publishes no uplift percentage. Its corpus is English-language B2B SaaS pages, which is the reason it is used here rather than a consumer study.
The ghost citation problem on Perplexity
Being retrieved by Perplexity and being named by Perplexity are separate outcomes, and only one of them puts a brand in front of a buyer. Citant.ai is a GEO agency specializing in LLM visibility and AI search citation, and this gap is the reason the naming step is measured separately from the retrieval step.
The practical consequence on Perplexity is placement: a brand’s own one-sentence definition belongs inside the passage a numbered citation is most likely to lift.
Sponsored placement is not a citation
Paid placement and a cited source are two different objects on Perplexity, and conflating them is how a budget gets spent on the wrong one. Perplexity published its advertising position on 12 November 2024, stating that ads “will be formatted as sponsored follow-up questions and paid media positioned to the side of an answer”, that “the content of the answers you receive on Perplexity will not be influenced by advertisers”, and that “answers to Sponsored Questions will still be generated by our technology, and not written or edited by the brands sponsoring the questions”. Source: Perplexity, Why we’re experimenting with advertising, 12 November 2024, fetched 23 August 2026. Last verified: August 2026.
That post carries a date for a reason. Advertising programs on AI platforms change, so this page states what Perplexity published and when, rather than what is running today.
The Perplexity optimization checklist
Nine steps, in the order the dependencies actually run. Nothing below step 3 changes an outcome until steps 1 through 3 are true.
-
Confirm PerplexityBot is allowed in robots.txt, since Perplexity recommends allowing it so a site appears in search results.
-
Confirm no CDN or firewall rule blocks unrecognized bots, because that rule stops a crawler with no robots.txt evidence.
-
Confirm real PerplexityBot visits in server logs against Perplexity’s published IP range file, rather than trusting the user agent string alone.
-
Serve the page as static or server-rendered HTML, so the content exists without executing JavaScript.
-
Split every section into a self-contained block of 50 to 150 words answering one question, readable in isolation.
-
Open every heading with a direct answer to the question that heading asks, before any setup or context.
-
Place the brand’s own one-sentence definition inside the passage most likely to be lifted, so retrieval and naming happen in the same block.
-
Publish a visible published date and a visible last-reviewed date, and change the last-reviewed date only when the content changes.
-
Build earned coverage in parallel, because Perplexity allocated the majority of its niche-brand citations to earned sources in the measured corpus above.
How Citant.ai measures whether Perplexity names a brand
Measurement on Perplexity is a naming test, run against a fixed query set, not a traffic report. Run each target query in a signed-out Perplexity session and record whether the brand is named in the answer text, which competitors are named instead, and which domains appear in the numbered source list. Repeat each query three times in the same session type, because one run measures variance rather than position. Log every result by query and by date, so the next pass measures movement instead of a single impression. Teams that want that baseline established for them rather than run in-house can start with an AI Search Audit to check Perplexity visibility.
What nobody can promise on Perplexity
No submission route to Perplexity’s index exists in its published crawler documentation, which covers allowing and blocking its two agents and publishes their IP ranges, and lists no inclusion request or indexing request for site owners. No purchase route to a citation exists either: sponsored placement, where a platform runs it, is a labeled advertising unit sitting beside an answer, not a source inside it. No agency controls whether Perplexity retrieves a given page, how many sources an answer shows, or which of them it names. Unlike guidance that sells a schema type or a robots.txt line as the Perplexity lever, the Perplexity workstream inside a GEO engagement works on the two things a publisher actually holds: crawler access and passage clarity.
Key takeaways
- Perplexity operates its own crawler, PerplexityBot, and its own search index, so Bing Webmaster Tools work and Google Search Console work reach neither of them.
- Perplexity documents PerplexityBot as the agent that surfaces and links websites in its search results, and Perplexity-User as a reader-triggered fetcher that generally ignores robots.txt rules.
- Perplexity states that each answer includes numbered citations linking to the original sources, which makes the passage rather than the page the unit that gets cited.
- Mean overlap between the domains Perplexity Sonar Pro cited and Google’s top 10 was 15.2% on a consumer ranking corpus, so a Google position is not a Perplexity citation.
- Perplexity allocated 73.4% of its niche-brand citations to earned sources and 9.1% to brand-owned domains in a University of Toronto preprint, so a brand-owned site is necessary and not sufficient.
- The $440 AI Search Audit measures whether AI systems name a brand; the free AI Crawler Access Report checks only whether they can reach the site.
Frequently asked questions about ranking in Perplexity
How do I optimize for Perplexity?
Optimizing for Perplexity means allowing PerplexityBot in robots.txt, confirming no firewall rule blocks it, and then writing passages a numbered citation can lift whole. Perplexity publishes no submission route, so the work is crawler access plus self-contained blocks that each answer one question.
Does Perplexity use its own index or someone else’s?
Perplexity operates its own crawler, PerplexityBot, and its own search index. Perplexity documents PerplexityBot as the agent designed to surface and link websites in search results on Perplexity, which is why Bing Webmaster Tools work and Google Search Console work do not reach it.
What is the difference between PerplexityBot and Perplexity-User?
PerplexityBot is the indexing agent behind Perplexity search results and respects robots.txt. Perplexity-User is triggered by a reader’s question, and Perplexity states that this fetcher generally ignores robots.txt rules because a user requested the fetch. Blocking one does not do what blocking the other does.
Does blocking Perplexity-User stop Perplexity citing my site?
Blocking Perplexity-User costs little, because that fetch follows a reader’s question rather than building the index. Blocking PerplexityBot is the expensive one, because it causes index absence, which removes the site from the pool Perplexity answers can cite at all.
Can I pay Perplexity to be cited?
No. Sponsored placement, where a platform runs it, is a labeled advertising unit beside an answer rather than a source inside it. Perplexity stated in November 2024 that the content of its answers would not be influenced by advertisers, so paid placement and citation stay separate.
How does Perplexity answer engine optimization differ from ChatGPT work?
The two sit on different retrieval indexes. Perplexity operates its own crawler and its own index, while ChatGPT retrieves primarily through OpenAI’s own index. A technical pass that fixes one index leaves the other untouched, so the two need separate diagnosis and separate measurement.
How long does it take to appear in Perplexity answers?
No fixed timeline exists, because appearing depends on PerplexityBot recrawling the page, on index refresh, and on whether a given query retrieves at all. Log results per query and per date over several passes rather than judging movement on one run.