
Perplexity SEO means making sure PerplexityBot can crawl your pages, then writing pages that answer a question in a self-contained passage Perplexity can quote and link. Keep content current and clearly dated, state facts precisely, and check your citations by running real prompts, because Perplexity does not publish a ranking formula you can optimise against.
AI answer tools are still a minority habit for news, but a growing one: the Reuters Institute Digital News Report 2025 found that 7% of people use AI chatbots for news each week, rising to 15% among under 25s.
Perplexity is useful to study because it behaves like a search engine that writes. You ask a question, it searches the web, and it answers with numbered citations to the pages it used. That visible citation habit makes it one of the easier AI products to test. This guide explains what Perplexity documents about its crawlers, how I would structure a page so a passage can be quoted, and how to check whether you are being cited. It is part of my wider work on generative engine optimisation.
How Perplexity finds and cites pages
Perplexity’s help centre describes the process in three steps: it interprets your question, searches the internet for information from sources such as articles, websites and journals, and compiles the most relevant insights into an answer. Its page on how Perplexity works also says each answer includes numbered citations linking to the original sources.
That has two practical consequences. First, retrieval happens at question time, so a page does not have to be old and authoritative to be considered, though it does have to be findable and readable. Second, the answer is built from passages. The number next to a sentence points to a page, but the sentence itself was drawn from one specific stretch of text. If your page has no stretch of text that stands alone as a good answer, it is hard to cite.
I should be honest about the limit here. Perplexity does not publish how it ranks or selects sources, so I will not tell you it favours a particular domain type or content format. What I can describe is what is documented and what I see when I test. For prompts in your own field, the quickest way to learn what gets cited is to read the citations yourself, which I come back to below.
Step 1: Allow PerplexityBot and understand the other agent
Perplexity’s crawler documentation names two user agents, and they behave differently.
| User agent | Purpose according to Perplexity | robots.txt behaviour |
|---|---|---|
PerplexityBot |
Surface and link websites in Perplexity search results. Not used to train foundation models. | Respects robots.txt. Perplexity says to allow it if you want visibility in its results. |
Perplexity-User |
Supports user-initiated requests, where it may visit a page to answer a question and include a citation. | Perplexity says it generally ignores robots.txt rules because a person triggered the visit. |
The decision rule is straightforward. If you want to be cited in Perplexity results, do not disallow PerplexityBot. If you want to stop Perplexity visiting a page on a user’s behalf, robots.txt is not the control for that, according to Perplexity’s own description, so you would need server side measures. Perplexity also publishes IP ranges for both agents in JSON files linked from that documentation, which matters in the next step.
Check your firewall as well as robots.txt
This is where I spend most of my time on a new site. Open robots.txt and confirm there is no blanket disallow. Then check your CDN, firewall and security plugins. Many bot protection tools block anything unfamiliar, and a crawler that receives a challenge page never sees your content. Use the published IP ranges to build an allow rule or to verify that a request claiming to be PerplexityBot genuinely is. My guide to AI crawlers and robots.txt covers verification and the decision framework for every major agent, and my robots.txt guide covers the syntax.
Step 2: Write passages that can be quoted
Because the answer is assembled from fragments, I think of each section as a candidate quote. A good candidate has a heading that asks the question, a first sentence that answers it, and a few sentences of support that do not depend on the rest of the page.
A before and after example
This is an illustrative example, not a client case. Imagine a dental clinic’s page with the heading “Our approach to teeth whitening” and a first paragraph about the clinic’s philosophy and warm welcome. A passage like that is hard to quote because it does not answer anything. Now imagine the heading “How long does professional teeth whitening last?” followed by a first sentence stating the typical range the clinic’s dentists give and the factors that shorten it, then a short list of those factors. The second version can be lifted whole and still make sense.
Apply this to your own pages with a short checklist:
- List the five questions customers ask most before buying.
- Give each one its own heading in the customer’s wording.
- Answer in the first sentence, with the number, name or condition that makes it useful.
- Add one or two supporting sentences and, where it helps, a table or list.
- Remove references such as “as mentioned above” that tie the passage to other text.
Pages built like this also tend to perform in Google’s AI features, which is why I keep my approach to AI Overviews and Perplexity broadly aligned.
Step 3: Keep pages fresh and honestly dated
Perplexity’s Search API documentation shows results carrying both a publication date and a last updated date, which suggests that date information is something the system can read. I would not read more into that than it says, because the API is not the consumer product. But the lesson is cheap to apply, and it is good practice anyway.
- Show a visible publication date and a genuine last updated date on pages where recency matters, such as pricing, regulations, comparisons and statistics.
- Update the content when you change the date. A new date over old facts is a trust problem, not an optimisation.
- Replace stale numbers and name the year for any figure that will date.
- Retire pages that are wrong instead of leaving them to be quoted.
Step 4: Make your facts easy to verify
An answer engine that cites sources needs pages it can trust enough to point at. You can help by being specific and checkable. State who wrote the page and why they are qualified, link to primary sources for claims that matter, and keep names, addresses and service descriptions consistent with your other profiles. The reasoning is the same as in my guide to entity based SEO: the clearer the picture of who you are, the less guessing is involved.
When I rebuilt saadrazaseo.com in October 2026, I added author information to every article and connected the structured data into one schema graph covering Person, ProfessionalService, Service, BlogPosting and FAQPage. I did that so any reader, human or machine, gets the same description of who stands behind the advice. I make no claim that it changes Perplexity’s output, because I have not tested it in a controlled way and the platform does not document such an effect.
Step 5: Check your citations
Perplexity makes citation checking easier than most AI tools, because the sources are listed next to the answer. A repeatable routine takes about an hour.
- Write 15 to 20 prompts across category questions, comparisons and brand questions that match how your customers speak.
- Run each one and copy the cited domains into a spreadsheet, with a column for whether you appear.
- Open the cited pages for the prompts where you are absent. Note what they offer that you do not: a direct answer, a table, a recent date, a primary source.
- Repeat on different days. Answers vary, so only count patterns that recur.
- Fix the gaps, then re-run in a month.
The full sampling method, including how to record mentions, sentiment and accuracy across several assistants, is in my guide to tracking brand visibility in AI assistants. If Perplexity sends visits, you can identify them in analytics by referrer, and I explain the report setup in measuring AI referral traffic in GA4.
Reading the results sensibly
If competitors are cited and you are not, the cause is usually one of three things: your pages are not reachable, your pages do not answer the question directly, or third-party sources describe the competitor more often than you. The first is a technical fix, the second is an editing job, and the third is slower reputation work. Start with the first because it is the cheapest.
Perplexity SEO compared with Google SEO
| Area | Classic Google SEO | Perplexity |
|---|---|---|
| Output | A ranked list of links | A written answer with numbered citations |
| Unit of competition | The page | The passage |
| Crawler control | Googlebot and robots.txt | PerplexityBot respects robots.txt; Perplexity-User is user-triggered |
| Ranking transparency | Partial, through guidance | No published formula |
| Measurement | Search Console | Manual prompt sampling and analytics referrers |
The overlap is large. Technically sound pages, clear writing and a trustworthy brand help in both. If your fundamentals are weak, an audit from a technical SEO service will do more for Perplexity visibility than any tactic specific to it.
Questions about Perplexity SEO
How do I allow PerplexityBot?
Make sure your robots.txt file does not disallow the PerplexityBot user agent for the pages you want cited, and confirm your firewall does not challenge it. Perplexity publishes the IP ranges for the bot, so you can use them to allow or verify requests.
Does Perplexity use my content to train models?
Perplexity’s documentation says PerplexityBot is designed to surface and link websites in search results and is not used to train foundation models. Read the current page before you decide, since policies can change.
Why does Perplexity cite my competitor and not me?
Perplexity does not publish its selection logic, so the honest answer is that you have to test. Compare the cited pages with yours for direct answers, freshness, clear sourcing and third-party mentions, then fix the largest gap first.
Can I block Perplexity-User with robots.txt?
Perplexity says Perplexity-User generally ignores robots.txt rules because requests are triggered by a person. If you need to stop it, use server level controls and weigh the cost of making your pages unavailable to people who ask about you.
If you want a second pair of eyes on your crawler access and page structure, I offer a free audit. Send me your site and I will tell you what is blocking citations, and what you can safely ignore.
More AI search guides: start with the generative engine optimization guide, then go deeper:
- ChatGPT Search SEO: How to Get Your Business Cited
- Google AI Mode SEO: What It Changes and What to Do
- AI Visibility Tracking: Monitor Your Brand in AI Answers
- AI Referral Traffic in GA4: How to Measure It Properly
- llms.txt Explained: What It Is and Do You Need One
- AI Crawlers Explained: GPTBot, ClaudeBot, PerplexityBot
- Agentic SEO: How to Make Your Website Work for AI Agents
- Bing SEO for AI Search: Webmaster Tools, IndexNow and Copilot
- Reddit SEO: How Community Content Shapes Google and AI Answers
