How to get cited by Perplexity: become a source it favours
Allow PerplexityBot, answer questions where Perplexity looks (Reddit and niche directories), keep pages fresh and dated, and write self-contained answers.
LLaunchScaler·Published ·8 min read
To get cited by Perplexity, first make sure PerplexityBot can crawl your site, then show up where Perplexity looks: Reddit threads, the niche directories of your category, and recently updated pages. Put a visible date and matching datePublished and dateModified on your content, and write each answer as a self-contained passage under a clear question heading.
Perplexity links the pages it used in its answers, so you can see which sources it picked for your buyers' questions and work toward being one of them.
How does Perplexity choose which sources to cite?
Perplexity finds pages two ways: PerplexityBot, which it says is designed to surface and link websites in its search results, and Perplexity-User, which fetches a page live when a question needs it. It then cites the pages it used. Two independent studies show which kinds of pages those tend to be: community discussions, especially Reddit, and specialised directories for the topic.
Profound analyzed 680 million AI citations from August 2024 to June 2025. Reddit was Perplexity's most-cited domain, with 6.6% of all its citations, and 46.7% of the share among its ten most-cited sources. The rest of that top ten was YouTube (13.9%), Gartner (7.0%), Yelp (5.8%), LinkedIn (5.3%), Forbes (5.0%), NerdWallet (4.5%), TripAdvisor (4.1%), G2 (4.0%) and PCMag (3.7%).
Yext analyzed 6.8 million citations from location-based consumer questions across Gemini, ChatGPT and Perplexity. It found Perplexity "consistently shows a strong preference for industry-specific directories": for unbranded, subjective questions, they made up 24% of its citations, the highest of the three engines.
What Perplexity cites
Evidence
What it means for you
Reddit threads
46.7% of top-ten share, 6.6% of all citations (Profound)
Questions, answered
What people ask about this
01
How does Perplexity choose which sources to cite?
Perplexity surfaces pages found by PerplexityBot, plus live fetches made when a user asks, and cites the pages it used. Studies of its citations show a strong lean toward Reddit, YouTube and category-specific directories, and toward recently updated pages.
Be part of the discussions in your category's subreddits
Industry directories and review sites
24% of citations for unbranded subjective queries (Yext); G2, Gartner and TripAdvisor in the top ten (Profound)
Get listed, accurately, where your category is reviewed
YouTube
13.9% of top-ten share (Profound)
Video explanations of the problems you solve
Recently updated pages
Median source age 32.5 days for SaaS and tech queries (AI+Automation)
Update and date your key pages
Step 1: Can PerplexityBot reach your pages?
Allow PerplexityBot in robots.txt and let its requests through your firewall. Perplexity's crawler documentation says PerplexityBot "is designed to surface and link websites in search results on Perplexity," recommends allowing it and its published IP ranges "to ensure your site appears in search results," and says it is "not used to crawl content for AI foundation models."
A robots.txt group that keeps Perplexity able to cite you:
A named group replaces the * group for that bot, so repeat any Disallow lines you still want. Perplexity says changes can take up to 24 hours to take effect.
The second agent behaves differently. Perplexity-User visits a page when a user's question needs it; Perplexity says it "is not used for web crawling or to collect content for training AI foundation models," and that "since a user requested the fetch, this fetcher generally ignores robots.txt rules."
Firewalls can block Perplexity silently. Perplexity's documentation gives WAF instructions for Cloudflare and AWS: allow requests that match both the user agent (PerplexityBot or Perplexity-User) and the IP ranges it publishes at perplexity.com/perplexitybot.json and perplexity.com/perplexity-user.json. On Cloudflare, check your AI bot policies too: from September 15, 2026, new domains block bots classified as Agent, which Cloudflare defines to include chat fetch bots, on pages that display ads. The PerplexityBot robots.txt guide covers the full setup.
Step 2: How do you show up on Reddit without spamming?
Answer real questions in the subreddits where your buyers ask them, with specific, useful replies, and say who you are when you mention your own product. Reddit's rules point the same way: Rule 2 asks users to "participate authentically in communities where you have a personal interest," and not to spam or manipulate content, and another rule tells users not to intentionally mislead others. Each subreddit's own rules are enforced by its moderators, so posts that break them can come down.
What that looks like in practice:
Find the threads Perplexity already cites for your category: ask Perplexity your buyers' questions and note every reddit.com source in the answers.
Read each subreddit's own rules before posting; many limit self-promotion to specific threads or days.
Reply where you have a real answer. Explain the method, the trade-offs and when your product is the wrong choice.
Disclose your affiliation when you mention your product, and link only when it answers the question.
Build a history of helpful answers that don't mention your product at all.
Step 3: Which directories and review sites should list you?
The ones Perplexity already cites for your category's questions. Yext's summary of its findings puts it plainly: "Perplexity rewards specialization," and being present and accurate in trusted niche directories signals authority. For a software product, that means the review sites and directories buyers in your category use, more than general business listings.
Find and fix them in one pass:
Ask Perplexity five of your buyers' questions and list every directory or review site in the sources.
Check whether you have a listing on each, and claim or create it.
Make the name, one-line description, category, pricing and link identical across all of them and your own site.
Ask real customers for reviews on the one or two sites that appear most often.
Step 4: How fresh does content need to be?
Fresher than for Google. An AI+Automation study compared the age of the top three cited sources on Perplexity and Google. For medium-velocity topics (SaaS reviews, tech comparisons, e-commerce), Perplexity's median was 32.5 days and Google's 108.2 days, about 3.3 times fresher. For fast-moving news it was 1.8 days against 28.6.
The study is small and independent, so read it as direction, not a rule. The practical takeaway is to review your most important comparison, pricing and how-to pages on a schedule, and update them when the facts change: new prices, new features, new competitors, new data. The same study warns that changing the displayed date without real changes is counterproductive, and Google's own byline guidance asks for dates that are accurate and consistent.
Step 5: How do you mark dates so Perplexity can read them?
Show a labelled date on the page, such as "Last updated: September 2026", and put the same dates in your structured data as datePublished and dateModified in ISO 8601 format with a timezone. The AI+Automation study reports that in its testing, pages without parseable dates were effectively treated as undated.
Google's Article structured data documentation describes both properties, and its byline guidance asks you to label dates clearly ("Published", "Last updated"), keep the visible and structured dates consistent, and never use future dates. A minimal block:
<script type="application/ld+json">
{
"@context": "https://schema.org",
"@type": "Article",
"headline": "How to choose a CRM for a five-person sales team",
"datePublished": "2026-03-02T09:00:00+00:00",
"dateModified": "2026-09-14T11:30:00+00:00",
"author": { "@type": "Person", "name": "Jane Doe" }
}
</script>
Keep <lastmod> in your sitemap equal to the same modified date, so every date signal a crawler finds agrees.
Step 6: How should you write pages Perplexity can quote?
Put each buyer question in a heading and answer it in the first 40 to 60 words beneath, so the passage makes sense lifted out on its own. Add figures with their sources, use tables for comparisons, and keep the content in the server HTML: Vercel's December 2024 study found PerplexityBot doesn't render JavaScript.
The GEO study by Aggarwal and colleagues tested content changes on Perplexity directly and measured visibility improvements of up to 37% there. Across its benchmark, the strongest changes were adding statistics, quotations and cited sources, while keyword stuffing offered "little to no improvement."
A page Perplexity can use has:
one question per heading, phrased the way a buyer asks it
a direct answer first, then the detail
numbers with a linked source for each
a comparison table where options differ
a visible, accurate "Last updated" date
How do you check whether Perplexity cites you?
Ask Perplexity 10 questions a buyer in your category would ask, not your brand name, and record for each answer whether your site is cited, whether your brand is named, and which sources are cited instead. Repeat monthly with the same wording. The sources list tells you exactly which Reddit threads, directories and articles you need to be part of.
Track the cited sources as carefully as your own result. If the same three Reddit threads and two directories appear across your questions, those five pages are your plan for the month. The measuring AI visibility guide gives a tracking sheet that covers Perplexity alongside the other engines.
See whether Perplexity and six other engines cite you
The technical half is free to check. Run the free scan first, then open the full audit on your report. The free scan takes just your address, no account, and its 156 checks include whether robots.txt or your firewall blocks PerplexityBot, OAI-SearchBot or Claude-SearchBot, whether your main content is in the server HTML, and whether your sitemap carries lastmod dates.
The full audit is a one-time $19 for your domain, and it never expires. It puts 10 questions buyers in your category ask to 7 answer engines, Perplexity included, reads all 70 answers and shows which ones cite you. It also adds the checks that need those answers or outside data: your Reddit presence, your presence on third-party review and directory sites, on-page freshness signals, and the competitors the engines name instead of you.
Yes, if you want to be cited. Perplexity recommends allowing PerplexityBot and its published IP ranges to ensure your site appears in its results, and says PerplexityBot is not used to crawl content for AI foundation models.
03
What is Perplexity-User?
It is the fetcher Perplexity uses when a user's question needs a page visited. Perplexity says it is not used for web crawling or model training, and because a user requested the fetch, it generally ignores robots.txt rules.
04
Does Perplexity cite Reddit a lot?
Yes. Profound's analysis of citations from August 2024 to June 2025 found Reddit held 46.7% of the share among Perplexity's ten most-cited sources, and 6.6% of all Perplexity citations, more than any other single domain.
05
How fresh does content need to be for Perplexity?
A study by AI+Automation found the median age of Perplexity's top three cited sources was 32.5 days for SaaS, tech and e-commerce queries, against 108.2 days for Google. Update pages when facts change, and show the date.