TGViewer
Artificial Intelligence AI News Artificial Intelligence AI News @machinelearningresearchnews · 3.49K subscribers
Post #1556 633
Search is now the bottleneck for AI agents, and Perplexity just rebuilt theirs from the ground up.

Photon is their new retrieval and ranking engine, written in Rust. It replaces the open-source engine they had forked for years.

The old engine's problem: the index outgrew RAM. Cold reads caused page faults, p99 sat near 800 ms, and index merges pushed it to ~1.2 s.

Photon's fix, in 4 moves:

1/ Compact posting lists that pick a format per block: inline, sparse arrays, or bitmaps
2/ A budgeted, WAND-like traversal that skips candidates before reading exact scores
3/ Disk reads batched through io_uring, so waits overlap instead of piling up
4/ Index builds moved off the serving nodes entirely

The result in production: p99 dropped to ~65 ms, on ~20% fewer machines, while storing ~2.5x more data per document.

For developers, Photon powers a new Fast Search mode:

→ 160 ms p50 / 230 ms p95 per call
→ $1 per 1K requests (vs $5 standard)
→ ~68% lower estimated agent task cost at comparable quality

Technical details + full breakdown in the reply 👇
Full breakdown: https://marktechpost.com/2026/09/30/perplexity-introduces-photon-a-rust-based-retrieval-engine-that-cuts-p99-latency-from-800-ms-to-65-ms/

Technical details: https://perplexity.ai/hub/blog/photon
  • 👍 3
  • 🤣 1
More from @machinelearningresearchnews
  1. Oct 10, 2026Microsoft just released Microsoft-Decision-1, a decision-scoring model that returns calibr…
  2. Oct 7, 2026Meta just open-sourced Rebalancer, the assignment solver that has run resource allocation…
  3. Oct 6, 2026Mistral AI Releases Mistral Large 4 (Le Chonk): A 1.05T Parameter Multimodal MoE Model Mis…
  4. Oct 5, 2026[AI Model Family Series #1] We just published the complete story of Alibaba's Qwen: every…
  5. Oct 2, 2026Cloudflare Releases Clef: Open-Weight Decision Models That Return Typed Probabilities Inst…
  6. Sep 30, 2026Google DeepMind just announced Gemini 4 Argon, raising the output limit from 64K to 1M tok…
Threads Profile ViewerView any public Threads profile without an account.Open ThreadLook →Writing with AI? Make it sound human.Metric37 rewrites AI drafts so they read naturally. Free AI detector, 1,500 words free.Try Metric37 →