Aside: AI memory on 80K+ devicesHow Aside powers AI memory across 80,000+ devices · 113M documents · 24.8B tokens · 150+ countries

Read moreRead the case study
Moss
usemossStart Free

On-Device

On-Device and Edge Semantic Search

Semantic search that runs where your app runs. Built in Rust and compiled to WebAssembly for browsers, mobile, desktop, and edge. No cloud infrastructure required. Works offline.

Get StartedTalk to an Engineer
offline notes

> Query: Tuesday standup notes

Searching local index...3ms

Top match found

Tuesday Standup — July 28

  • · Resolve sync blockers before release
  • · Deployment freeze begins Thursday
  • · Review Q3 roadmap changes with product

<10ms

on-device retrieval

100%

offline capable

0

cloud infrastructure required

Semantic search that runs where your app runs.

Built in Rust and compiled to WebAssembly for browsers, mobile, desktop, and edge. No cloud infrastructure required. Works offline.

The Problem

Not every search query should leave the device

Cloud vector databases assume every query travels over the network. But for mobile apps, desktop tools, browser extensions, and edge deployments, that assumption adds latency, cost, and privacy risk. Users on slow connections wait seconds for results. Users on no connection get nothing at all. And every query sent to a cloud service is a data point you no longer control. On-device search solves all three problems at once.

Cloud vector databases assume every query travels over the network. But for mobile apps, desktop tools, browser extensions, and edge deployments, that assumption adds latency, cost, and privacy risk. Users on slow connections wait seconds for results. Users on no connection get nothing at all. And every query sent to a cloud service is a data point you no longer control. On-device search solves all three problems at once.

How Moss Solves This

1

Runs everywhere via WebAssembly

Moss compiles to WASM and runs in browsers, Electron apps, React Native, mobile WebViews, and edge servers. One runtime for every deployment target.

2

Offline-first architecture

Once loaded, the index lives in-memory. All queries execute locally with zero network dependency. Data syncs automatically when connectivity returns.

3

Lightweight runtime

The Moss WASM runtime is compact enough for mobile and edge deployments. No heavy dependencies, no GPU requirements, no infrastructure to manage.

Ship real-time retrieval in minutes

Explore docs
1
2
3
4
5
6
7
from moss import MossClient

client = MossClient(PROJECT_ID, PROJECT_KEY)

index = await client.load_index("product-catalog")

results = await index.search("wireless headphones")

Frequently asked questions

Ready to ship faster AI products?

Moss gives you production-ready semantic retrieval without infrastructure complexity.

Test performanceTalk to an Engineer
Moss
AICPA SOC 2 Type 2HIPAA

Product

Founding AgentLocal Search

Use Cases

Voice AIAI CopilotsIn-App SearchOn-Device AI

Company

PricingBlogCareersBrand Kit

Resources

DocsGlossaryBenchmarks

Integrations

DSPyElevenLabsLangChainLiveKitMCP ServerNext.jsPipecatVAPIVercel AI SDKVitePress

© 2026 MOSS

Privacy PolicyTerms of ServiceTrust Center