Offline · Step 1 of 12
Crawl the web
Exa’s crawler keeps discovering new URLs and fetches them across a distributed network of machines and IP addresses.
Source: Exa blog, Mar 2025
Neural web search, end to end
Exa runs two systems that never meet in real time. An offline pipeline turns the web into compact vectors and files them into clusters. An online path turns your query into a vector and scores only the nearest clusters. Press play, or pick a step. Click any box in the diagram to jump to it.
Offline · Step 1 of 12
Exa’s crawler keeps discovering new URLs and fetches them across a distributed network of machines and IP addresses.
Source: Exa blog, Mar 2025
The same page at each stage of compression. Squares are drawn with area to scale.
Per billion pages: 8 TB, then 512 GB, then 32 GB. Exa says the full index uses less memory than a gaming PC; the full vectors stay on disk for reranking.
How Claude Code on this machine reaches Exa.
mcp.exa.ai/mcp, registered in ~/.claude.json with your API key.web_search_exa (this diagram), web_fetch_exa (read a URL from Exa’s copy), agent_run.E:\CLAUDE.md §3 already says entity and “things like X” questions go to Exa. Plain lookups go to Firecrawl, and single URLs go to Crawl4AI.