Aside: AI memory on 80K+ devicesHow Aside powers AI memory across 80,000+ devices · 113M documents · 24.8B tokens · 150+ countries

Read moreRead the case study
Moss
usemossStart Free

Voice AI

Semantic Search for Voice Agents

From voice agents to AI copilots, Moss gives your product <10ms semantic retrieval without managing vector infrastructure.

Start FreeExplore Docs

I need to reschedule my appointment

Sure! I see you're booked with Dr. Chen, Thursday at 2pm. What day works better for you?

Does next Tuesday work?

Yes, Tuesday at 2pm is open. Want me to move it there?

Yes, please.

✓Rescheduled to Tuesday, 2pm

<10 ms

P99 retrieval latency

0 ms

network round-trips

100x

faster than cloud vector DBs

Your voice agent needs context in under 10 milliseconds.

Network round-trips to cloud vector databases add 300-900ms of dead air. Moss runs retrieval locally, inside your agent runtime, so your agent responds without the pause.

The Problem

Dead air kills voice agents

Traditional RAG sends retrieval over the network to Pinecone or Qdrant — and round-trips mean silence. Users notice gaps as short as 200ms. The LLM was never the bottleneck. Retrieval is.

How Moss Solves This

1

Local retrieval, zero network hops

Moss runs search inside your agent runtime. No network round-trip to a cloud database. Retrieval completes in under 10ms, locally.

2

Built for streaming conversation

Designed for the real-time loop of ASR, retrieval, LLM, and TTS. Moss fits into the critical path without adding perceptible latency.

3

Works with every voice stack

Drop-in integration with LiveKit, Pipecat, VAPI, ElevenLabs, and Hume AI. Install the SDK and start querying in minutes.

Ship real-time retrieval in minutes

Explore docs
1
2
3
4
5
6
7
from moss import MossClient

client = MossClient(PROJECT_ID, PROJECT_KEY)

docs = [{ "text": "How do I track my order?" }]

await client.add_docs("my-index", docs)

Frequently asked questions

Ready to ship faster AI products?

Moss gives you production-ready semantic retrieval without infrastructure complexity.

Test performanceTalk to an Engineer
Moss
AICPA SOC 2 Type 2HIPAA

Product

Founding AgentLocal Search

Use Cases

Voice AIAI CopilotsIn-App SearchOn-Device AI

Company

PricingBlogCareersBrand Kit

Resources

DocsGlossaryBenchmarks

Integrations

DSPyElevenLabsLangChainLiveKitMCP ServerNext.jsPipecatVAPIVercel AI SDKVitePress

© 2026 MOSS

Privacy PolicyTerms of ServiceTrust Center