← Policy library/Perplexity AI: Citation Guidelines & Crawler Optimization
Policy library

Perplexity AI: Citation Guidelines & Crawler Optimization

Technical and architectural specifications for ensuring your website is crawled by PerplexityBot and cited in Perplexity answers.

AI Summary: Perplexity AI generates answers by retrieving web content via PerplexityBot and synthesizing direct responses with numeric source citations. Publishers optimize for Perplexity by allowing PerplexityBot in robots.txt, structuring content with high information density, and publishing an llms.txt guide.

The Perplexity Retrieval Engine Architecture

Perplexity operates as an answer engine that combines custom web search indexes with frontier LLMs. For nearly every prompt, Perplexity performs real-time retrieval across multiple sources, ranks the snippets for relevance and authority, and generates a structured synthesis with numbered inline citations.

Being cited in Perplexity drives exceptionally high-intent, qualified referral traffic because users click source cards to verify technical claims.

Crawler Directives for PerplexityBot

Perplexity operates under the user-agent token PerplexityBot:

configuration / code
User-agent: PerplexityBot
Allow: /
Disallow: /private/
Disallow: /api/

[!IMPORTANT] If your robots.txt has a blanket User-agent: * Disallow: /, PerplexityBot will honor the restriction and exclude your domain from live search citations. Always provide an explicit allow block for PerplexityBot.

4 Rules for Earning Perplexity Citations

1. High Information Density & Low Fluff

Perplexity's reranking algorithms prioritize content with high factual density. Eliminate lengthy introductions and marketing filler. State facts, methodologies, and outcomes directly.

2. Tabular Data & Numerical Specifics

Include verified data tables, metrics, and technical benchmarks. When Perplexity answers comparison queries ("What are the rate limits of X vs Y?"), it extracts data from well-structured tables.

3. Deploy an /llms.txt File

Perplexity supports the /llms.txt standard—a markdown file at your domain root that maps your site's canonical documentation. This provides a clean road map for AI retrieval.

4. Fast Edge Delivery & Clean SSR

Ensure critical documentation is rendered as static or server-rendered HTML. PerplexityBot prioritizes fast HTTP response times to meet its sub-second answer generation deadlines.


Audit your content architecture for Perplexity citation readiness. Run a comprehensive audit with Geolify.ai.

Related policies