AI Tools & Products

Perplexity Introduces Photon: A Rust-Based Retrieval Engine That Cuts p99 Latency From 800 ms to 65 ms

· September 30, 2026
Perplexity Introduces Photon: A Rust-Based Retrieval Engine That Cuts p99 Latency From 800 ms to 65 ms

What changed

Perplexity launched Photon, a new retrieval and ranking engine written in Rust. Photon replaces a fork of an open-source engine previously used in their AI-native search stack. It now handles all production retrieval and ranking tasks. Photon also powers a Fast Search mode in the Perplexity Search API, delivering single-call latencies around 160 milliseconds and improving p99 latency from 800 milliseconds to 65 milliseconds.

Why builders should care

Cutting p99 latency by nearly an order of magnitude matters for any product relying on fast, reliable search and ranking at scale. Lower latency improves user experience by speeding up responses, which is critical for AI applications requiring quick information retrieval. Using Rust likely boosts performance and reliability compared to existing engines implemented in other languages. Developers working on search, recommendation systems, or real-time AI inference can learn from this example of reengineering core infrastructure to gain clear operational advantages.

The practical takeaway

Photon shows that investing in a custom, high-performance retrieval engine can substantially lower latency bottlenecks in AI-driven search stacks. Builders facing high query volumes or slow response times should consider Rust-based implementations for mission-critical retrieval and ranking workflows. This approach can unlock faster API responses and support advanced features like Fast Search modes without sacrificing quality or scale.

What to watch next

Monitor whether Photon adoption spreads beyond Perplexity or sparks open-source equivalents in Rust. Expect competitors to focus more on reducing tail latency for search and ranking as user expectations for real-time AI grow. Also, watch for any public benchmarks or architectural write-ups from Perplexity that shed light on how Rust contributed to these gains. The broader trend should push AI infrastructure teams to re-evaluate language and design choices at the foundation of their search stacks.

AI Quick Briefs Editorial Desk

Stay ahead of AI Get the most important AI news delivered to your inbox — free.