Skip to main content
usedby

under 200ms

“Recall returns in under 200ms on average and under 800ms at P99”

As published on zilliz.com. Captured by usedby on Oct 6, 2026.

What happened

Plaud uses Zilliz Cloud as the retrieval layer behind its ContextOS system. It stores semantic vectors and scalar filter fields to support search across a user's recordings, Ask Plaud (RAG), and long-term personal memory for its agents.

Summary written by usedby from the source page, in English. The figures are those of Zilliz and Plaud, not ours.

  • under 800ms“under 800ms at P99”

Recall returns in under 200ms on average and under 800ms at P99, giving the end-to-end Ask AI pipeline room to stay responsive even as data grows into the billions.

From the page. zilliz.com, captured Oct 6, 2026

AI is moving from answering one-off questions to agents that remember, which makes AI memory the heart of consumer AI products. Zilliz Cloud gives us a solid foundation we can trust for agentic memory retrieval at a massive scale, so our team can put its energy into product innovation and user experience, not the plumbing beneath it.

Charles Liu, Co-founder & CTO, Plaud. Source, captured Oct 6, 2026

What the story claims, and what we checked

We compared the story with its live page on Oct 6, 2026.

  • The figure: under 200msCheckedPrinted word for word on the page, near the name of Plaud.
  • The passage quoted aboveCheckedCopied word for word from the page, near the name of Plaud.
  • Plaud uses ZillizCheckedConfirmed line. Latest check across sources: Oct 6, 2026.
  • The result itselfNot checkedWe quote it; we did not measure it.

Same company, same tool or same industry.

~45 msP99 dense retrieval latency in production acrossConsensus uses Zilliz. Another customer of Zilliz<200msNeural search latency with Exa Instant, reducedExa uses Zilliz. Another customer of Zilliz60-80%“60-80% of the time spent on consuming, digesting, and identifying relevant data points”Filevine uses Zilliz. Another customer of Zilliz