AINewsnow

Perplexity Introduces Photon: A Rust-Based Retrieval Engine That Cuts p99 Latency From 800 ms to 65 ms

Search is now the bottleneck for AI agents, and Perplexity just rebuilt theirs from the ground up. Photon is their new retrieval and ranking engine, written in Rust. It replaces the open-source engine they had forked for years. The old engine's problem: the index outgrew RAM. Cold reads caused page…

Read the full story at r/machinelearningnews ↗

Timeline · 1 report

  1. 2026-09-30 08:41 · r/machinelearningnews
    Perplexity Introduces Photon: A Rust-Based Retrieval Engine That Cuts p99 Latency From 800 ms to 65 ms

More stories

  1. Perplexity Releases pplx-embed-v2-context-9b-preview: A Contextual Embedding Model That Retrieves Answers and Their Supporting Evidence — MarkTechPost
  2. Clef: Open Weights decision model by Cloudflare — r/LocalLLaMA
  3. NVIDIA DGX Spark 64GB Gives Developers More Ways to Build and Scale Local AI — NVIDIA Blog
  4. Gemini 4 Argon: our next era of frontier intelligence — Google Gemini Blog
  5. Guided Vision in Gemini Live: built for accessibility — Google Gemini Blog
  6. Tavus unveils Griffin, the "first Human Interaction Model", which it says passed the "video Turing test", with 48% of users thinking it was human in live chats (@tavus) — Techmeme
  7. Google tests its plan for AI data centers in space with Project Suncatcher — Scientific American
  8. Google announces Gemini 4 Argon AI model, but you can't use it yet — Ars Technica AI

Get the daily brief of stories like this at 6:30 every morning →