SEO & AI Search
Perplexity SEO: How to Get Cited in Perplexity Answers, With Anobee's Own Citation Data (2026)
By Bibek Thapa · Updated · 13 min read
Quick Answer
Perplexity SEO is the work of getting a page cited as a source in Perplexity's generated answers. Three things matter most: Perplexity leans heavily on Google's top 10 (82% URL overlap in Semrush's 2025 study), its crawler must be able to reach you (check Cloudflare, not just robots.txt), and answer-first pages are easier to quote. Measure with a fixed prompt set, not rankings.

Table of ContentsOn this page
Key Takeaways
- Perplexity has two fetchers: PerplexityBot obeys robots.txt and feeds its search index; Perplexity-User fetches on demand and generally ignores robots.txt.
- Perplexity is the AI engine closest to Google: 82% URL overlap with Google's top 10 in Semrush's July 2025 study. Ranking in Google is most of the job.
- Since July 2025 Cloudflare blocks AI crawlers by default on new domains. On Cloudflare, that setting decides your Perplexity visibility before content does.
- In Anobee's fixed-prompt test, Perplexity cited anobee.com in all three runs but kept almost none of the same pages. Presence is stable; source choice is not.
- Perplexity sends referral traffic you can see in GA4 as perplexity.ai / referral, which ChatGPT often does not.
- The FETCH checklist (Fresh, Extractable, Trustworthy, Crawlable, Hyper-relevant answer first) is a pre-publish routine, not a ranking formula.
This is a refreshed version of Anobee's Perplexity SEO guide. The earlier version leaned on third-party ranking-factor percentages that could not be traced to a source, so they are gone. In their place: Perplexity's own crawler documentation, two Cloudflare reports that change the crawlability advice, and Semrush's overlap study. Also, something the first version did not have: three runs of Anobee's own fixed-prompt citation test, with Perplexity's actual source choices recorded.
What Perplexity SEO is
Perplexity does not show ten links. Instead, it runs a search, reads a set of pages, writes one answer and attaches source links to the sentences it took from them. There is no position two. A page is either in the source set or it is not. Therefore, the goal of Perplexity SEO is to be in that set for the questions your readers ask.
The audience is not small. Perplexity's CEO said the service handled 780 million queries in May 2025, about 30 million a day, growing more than 20% month over month [3]. However, Perplexity does not publish user counts. Later figures you see are third-party estimates; the May 2025 number is the last one stated by the company itself.
Two properties of Perplexity shape everything below. It fetches pages live rather than answering from a fixed training set. Also, it cites visibly, which means you can measure it.
How Perplexity chooses sources
Perplexity's retrieval is closer to Google's than any other AI engine's. Semrush's July 2025 study of 5,000 keywords found that Perplexity's cited sources had over 91% domain overlap and 82% URL overlap with Google's top 10 results [2]. By comparison, Google's own AI Mode managed roughly 54% domain and 35% URL overlap in the same study. ChatGPT showed the least alignment of all [2].

That single finding reorders the usual Perplexity SEO advice. If four out of five URLs Perplexity cites are already in Google's top 10 for the query, the main path into Perplexity's answer is the one you already know: rank in Google. Formatting, freshness and structure decide which of the retrieved pages get quoted. However, they rarely get a page retrieved that Google does not rank.
The earlier version of this article argued the opposite: that Google rankings do not guarantee Perplexity citations. That is still literally true, because a citation is never guaranteed. But the weight was wrong. Ranking is the first-order factor. Everything else is second-order.
What Perplexity says about its own crawlers
Perplexity documents two fetchers, and they behave differently [1]:
| Fetcher | What it does | robots.txt |
|---|---|---|
| PerplexityBot | "Designed to surface and link websites in search results on Perplexity." Not used to train foundation models. | Follows it. Allow it to be discoverable. |
| Perplexity-User | Visits a page when a user asks a question, to answer accurately and link to the page. | "Generally ignores robots.txt rules," because a user requested the fetch. |

Both publish IP ranges as JSON files. Perplexity says changes to access rules can take up to 24 hours to apply [1]. The practical reading: blocking PerplexityBot removes you from Perplexity's own index, which is where most citations come from. It does not stop a user-triggered fetch. Meanwhile, allowing it costs nothing, because the bot is not a training crawler.
What the studies can and cannot tell you
Third-party citation studies are useful for direction and unreliable for precision. Analyze AI's May 2026 analysis of 83,670 citations across ChatGPT, Claude and Perplexity found that pages in a question-and-answer or direct-answer format reached a top-3 citation slot 55% of the time, against a 31% average [6]. Freshness also correlated with placement [6]. However, it is a vendor study with its own tracking tool behind it. Treat the direction (answer-first pages do better) as credible and the exact percentages as that vendor's sample.
What no study has established is a weighted list of "Perplexity ranking factors". The earlier version of this guide published one. It has been removed, because the numbers could not be traced to any dataset.
What Anobee's own Perplexity SEO test shows
Since September 2026 Anobee has run the same prompt, "Search the web for Anobee and cite the sources you use", through ChatGPT, Gemini, Perplexity and Google AI Mode. Each Perplexity run used a fresh, signed-out session, and every source URL was recorded. The full method and datasets are in the AI Citation Index. The Perplexity column alone is instructive for Perplexity SEO.
| Run | Perplexity cited anobee.com | Distinct Anobee URLs | Retained from previous run | Notes |
|---|---|---|---|---|
| 9 Sep 2026 | Yes | 7 | — | Mixed in two unrelated lookalike domains |
| 14 Sep 2026 | Yes | 4 | 0 of 7 | Entity correct; source panel carried off-topic pages |
| 18 Sep 2026 | Yes | 11 | 1 of 4 (homepage) | 15-entry panel: 11 Anobee pages, one author site, three unrelated |

Four things stand out.
A small, new site gets cited. Anobee began publishing in mid-2026 and was in Perplexity's source set on all three runs. Domain age was not the barrier.
Presence is stable, source choice is not. Perplexity cited the site every time. Yet between the first two runs it kept none of its seven pages, and between the next two it kept only the homepage. If you track "were we cited" you will see a flat line. However, if you track which URLs, you will see churn.
Perplexity retrieves broadly and cites narrowly. On 18 September the source panel held 15 entries. Only five of the eleven Anobee pages were linked inline in the answer text. The panel is the retrieval set; the inline links are the citations a reader sees.
Noise is normal. Every run carried at least one unrelated page in the panel: a Yahoo article about power outages, a word list, a Reddit thread about watches. None reached the answer text. Moreover, Perplexity described Anobee correctly each time. Entity accuracy and source tidiness are separate properties.
The pages Perplexity chose on 18 September were the homepage, About, Contact, the Privacy Policy, the author page, and six articles. The entity pages, the ones that say who runs the site and how to reach it, kept appearing. That is a small sample. Nevertheless, it matches the advice below to make identity pages unambiguous.
The FETCH checklist for Perplexity SEO
FETCH is Anobee's pre-publish routine for pages that should be citable. It is a checklist, not a ranking formula. Also, it applies to Google's AI features as much as to Perplexity: Google itself says no special optimisation is needed for its AI features, and that ordinary SEO best practice continues to apply [7].
- F, Fresh. Update pages on fast-moving topics and show a real "last updated" date. Change the date only when the content changed.
- E, Extractable. Short headings that name the question, a direct answer under each, tables for comparisons, numbered steps for processes.
- T, Trustworthy. A named author with a linked bio, sourced numbers, and an About and Contact page that state who runs the site.
- C, Crawlable. PerplexityBot allowed, no Cloudflare or firewall rule blocking it, pages that render without JavaScript tricks.
- H, Hyper-relevant answer first. The direct answer to the page's main question in the first 100 words, in a sentence that makes sense on its own.

7 Perplexity SEO steps to get cited
Step 1: Make sure Perplexity can actually reach you
Start with robots.txt: open yourdomain.com/robots.txt and confirm there is no Disallow rule for PerplexityBot. The robots.txt guide for beginners covers the syntax.
Then check the layer most guides skip. Since 1 July 2025, Cloudflare asks every new domain whether to allow AI crawlers and blocks them by default. It also offers a one-click block for existing customers [5]. Consequently, a site behind Cloudflare with that setting on is invisible to PerplexityBot, no matter what robots.txt says. Log in to Cloudflare, open the AI crawl controls for the zone, and decide deliberately. Anobee runs on Cloudflare with an open robots.txt and no AI-crawler block, which is why Perplexity could cite it in the runs below; allowing search crawlers while blocking training crawlers is a reasonable middle setting if you want one.
One more thing to know. In August 2025 Cloudflare reported that when Perplexity's declared crawlers were blocked, undeclared crawlers using a generic Chrome user agent and rotating networks fetched the same content anyway [4]. Cloudflare began blocking that behaviour [4]. Two consequences follow. A robots.txt block is not a reliable way to keep content out of Perplexity. Conversely, if you want to be cited, an explicit allow is the only setting that makes your intent clear.
Verify: fetch your robots.txt, confirm PerplexityBot is not disallowed, and confirm your CDN or firewall is not challenging or blocking it.

Step 2: Rank in Google for the question
Given the 82% URL overlap with Google's top 10 [2], this is the step in Perplexity SEO with the most leverage and the least Perplexity-specific advice. Pick the question a page answers and check the live Google results for it. Then do the ordinary work: match the intent, cover what the top results cover, earn links. The on-page SEO checklist is the list Anobee runs. A page on Google's page three is unlikely to be in Perplexity's retrieval set, however well it is formatted.
Step 3: Answer the question in the first 100 words
Perplexity quotes passages. The passage most likely to be quoted is a self-contained answer near the top of the page, phrased close to the way the question is asked. Write it as one short paragraph, 40 to 60 words, that states the answer before the context. Anobee's CMS renders this as a "Quick answer" block. On WordPress, a bold lead paragraph does the same job.
Step 4: Structure the rest for extraction
Under each heading, put the answer to that heading's question in the first sentence. For Perplexity SEO, that first sentence is the one most likely to be quoted. Use a table wherever two or more things are compared, numbered lists for processes, and definitions stated as definitions ("X is …"). Avoid long paragraphs that bury the claim in the middle. Similarly, this is the advice Google gives for its own AI features, framed as helping readers rather than machines [7].
Step 5: Make the entity pages unambiguous
Perplexity kept citing Anobee's homepage, About, Contact and author page across runs. Those pages exist to say who is behind the site. Make sure they do: the same name everywhere, a real contact route, an author page with a bio, and consistent facts across the site and your profiles. The entity optimization guide covers the details. Also, if your brand name collides with another company's, as Anobee's does with Amobee, state the distinction on the About page yourself rather than leaving it to the model.
Step 6: Keep pages current, honestly
Freshness correlates with citation in every third-party study, including Analyze AI's [6]. However, the trap is faking it. Google's people-first content guidance lists changing a page's date to make it seem fresh, when the content has not substantially changed, among the signs of search-engine-first content [8]. Update pages when facts change, note what changed, and let the date follow the edit. For fast-moving topics like this one, that is roughly monthly. For a how-to that has not changed, it is never.
Step 7: Measure with referrals and a fixed prompt set
Perplexity, unlike ChatGPT, usually passes a referrer. In GA4, go to Reports, then Acquisition, then Traffic acquisition. Set the dimension to Session source / medium and search for "perplexity"; visits appear as perplexity.ai / referral. The GA4 setup guide covers the property side.
Referrals only capture clicks, and most citations do not produce one. Therefore, run a fixed set of ten to twenty prompts in a fresh Perplexity session each month. Record, per prompt, whether your site appears in the source panel, whether it is linked inline, and which URLs. The method for checking AI citations has the recording template. Keep the prompts identical between runs, because changing them breaks the comparison.
Where the earlier Perplexity SEO guide was wrong
Refreshing a page means saying what changed. Three things did.
First, the weighted "six ranking factors" table and figures such as "domain authority is 15% of Perplexity's scoring" had no traceable source and are gone. Second, the claim that Google rankings and Perplexity citations are loosely related understated the relationship. Semrush's overlap data says the opposite [2]. Third, the crawlability advice treated robots.txt as the whole story. Perplexity's own documentation and Cloudflare's reports show it is not [1][4][5].
What survived is the FETCH routine and the measurement approach, both now backed by Anobee's own recorded runs rather than by borrowed statistics.
Bottom line on Perplexity SEO
Perplexity SEO is mostly SEO. Rank in Google for the question, make sure PerplexityBot and your CDN are not standing in the way, put the answer at the top of the page, and keep your identity pages consistent. Then measure with a fixed prompt set, because a Perplexity citation is a snapshot. Anobee was cited in every run and kept almost none of the same pages between them. If ChatGPT is the next engine you want to be cited in, how to get cited by ChatGPT covers where its behaviour differs.
Frequently Asked Questions
What is Perplexity SEO?
Perplexity SEO is the practice of making a page eligible, retrievable and easy to cite in Perplexity's generated answers. Perplexity does not show a ranked list; it writes one answer and attaches source links. The work is being crawlable, ranking well in Google (which Perplexity's retrieval leans on), and stating answers clearly enough to be quoted.
Does Perplexity use Google rankings to choose sources?
Not officially, but its results track Google closely. Semrush's July 2025 study of 5,000 keywords found Perplexity's cited sources had 91% domain overlap and 82% URL overlap with Google's top 10, far higher than AI Mode or ChatGPT. Strong Google rankings are the most reliable path into Perplexity's source set.
Does blocking PerplexityBot in robots.txt keep my site out of Perplexity?
It keeps you out of PerplexityBot's index, which Perplexity says surfaces sites in its search results. It does not stop Perplexity-User, which fetches a page when a user asks a question and, according to Perplexity's documentation, generally ignores robots.txt. Cloudflare has also reported undeclared Perplexity crawlers evading blocks.
Can a new or small site get cited in Perplexity?
Yes. Anobee, a site that began publishing in mid-2026, was cited by Perplexity in all three runs of its fixed-prompt test in September 2026, with between 4 and 11 pages per run. The pages it chose changed almost completely between runs, so treat a citation as a snapshot, not a position you hold.
How do I track Perplexity traffic in Google Analytics?
In GA4, open Reports, then Acquisition, then Traffic acquisition, set the dimension to Session source / medium and search for perplexity. Perplexity passes perplexity.ai as the referrer, so visits appear as perplexity.ai / referral. For citations that never turn into clicks, run a fixed set of prompts monthly and record which pages appear.
Sources and References
- Perplexity — Perplexity Crawlers (developer documentation) ↩
- Semrush — How Google's AI Mode Compares to Traditional Search and Other LLMs (AI Mode Study) ↩
- TechCrunch — Perplexity received 780 million queries last month, CEO says ↩
- Cloudflare — Perplexity is using stealth, undeclared crawlers to evade website no-crawl directives ↩
- Cloudflare — Content Independence Day: no AI crawl without compensation ↩
- Analyze AI — Perplexity AI Ranking Guide: What Drives Citations ↩
- Google Search Central — Optimizing your website for generative AI features on Google Search ↩
- Google Search Central — Creating helpful, reliable, people-first content ↩
Was this guide helpful?
Your answer helps Anobee improve future updates.
Written by
Bibek Thapa
AI-Powered Digital Growth Strategist
Bibek Thapa works across AI workflows, SEO, AI search optimization, content strategy, website growth, and productivity systems. Anobee documents practical lessons, tools, experiments, and systems for improving digital presence.
- AI workflows
- Digital growth
- SEO
- GEO
- AEO
- Content strategy
- Website growth


