Exa
Web search API built for AI agents with sub-180ms latency and token-efficient highlights.
Exa is the practical pick for developers building AI agents that need live web data without burning tokens. The free tier is generous for prototyping, but watch per-request costs at scale. If your agent relies on real-time search, Exa is the top choice today.
Verified 8d ago · liveness 87/100 · cite: rightaichoice.com/tools/exa
- AI agents needing real-time web data for coding, research, or automation
- Developers building LLM-powered applications that require structured search results
- Lead generation and company enrichment for sales intelligence tools
- Market research applications needing up-to-date competitive data
- General-purpose web search for end-users with a graphical interface
- Projects requiring full HTML page retrieval without summarization
- Budget-constrained teams needing a free or low-cost search API
We scan live Reddit threads, YouTube comments, X posts, G2 reviews and other communities — and hand you an honest verdict in under a minute.
- Honest verdict, not marketing
- Real pros & cons from real users
- Attributed quotes with receipts
3 free scans · no card needed
Skip Exa if you need a free search API for high-volume or non-agentic workloads, or if your project requires full HTML retrieval without summarization.
Going past 10 results per request adds $1 per 1k additional results, which can inflate cost for bulk retrieval.
Exa's pay-as-you-go pricing suits startups and developers who need scalable search API without upfront costs. For high-volume use, consider dedicated enterprise plans, which may offer volume discounts. Compared to Perplexity or Brave APIs, Exa's pricing is competitive for agentic workloads due to token-efficient highlights reducing LLM costs.
In short
Exa — Web search API built for AI agents with sub-180ms latency and token-efficient highlights. Best for AI agents needing real-time web data for coding, research, or automation, Developers building LLM-powered applications that require structured search results, Lead generation and company enrichment for sales intelligence tools. Free to start; paid plans from $0.0121/mo.
Viability Score
How well maintained and how widely used is Exa? Built from what the vendor actually publishes (docs, changelog, tutorials, integrations, pricing), whether the site is live, and how much real users discuss it. How we calculate this
Last calculated: August 2026
How we score →Key Features
- Web search API with sub-180ms latency (Instant Search)
- Token-efficient highlights (up to 90% token reduction)
- Exa Agent API with natural-language queries and effort modes
- Deep Search for multi-step research with grounded citations
- Deep Reasoning type for deeper analysis (12-40s)
- Structured output extraction via output_schema (70M+ companies, 1B+ people)
- Scheduled Monitors with webhook delivery and deduplication
- Category-specific indexes: company, people, publication (350M), news
- Contents endpoint for full page contents with highlights
- Configurable latency presets: auto, instant, fast, deep-lite, deep, deep-reasoning
- Semantic search across code, docs, and web
- Exa Connect for live access to premium data providers
- Zero Data Retention option available
- SOC 2 Type II certified and single sign-on (SSO)
- MCP server support (including Exa Agent and Exa Connect)
About Exa
Exa is a web search API designed specifically for AI agents and LLM applications, offering a single endpoint for search, crawling, and deep research. It's the infrastructure behind coding agents like Devin and powers lead enrichment for HubSpot, delivering real-time web data in formats that models can use directly. With latency options from sub-180ms (Instant Search) up to 40 seconds for deep reasoning tasks, Exa balances speed and depth based on your agent's needs. The token-efficient highlights feature, powered by a specialized model, reduces token usage by up to 90%—passing over 25T tokens to models each week. Structured output extraction works across 70M+ companies and 1B+ people (with the publication index expanded to 350M), and the Exa Agent API handles natural-language queries with configurable effort modes (minimal to x-high). Deep Search runs multi-step agent workflows with grounded citations. Exa leads benchmarks on FRAMES (54.4% vs Perplexity's 44.5% and Brave's 21.6%), Tip-of-Tongue, and Seal0. Pricing is consumption-based with a free tier of $20 credits on sign-up and $10 monthly credits; Search is $7 per 1k requests, Contents $1 per 1k pages, Monitors $15 per 1k requests, and Deep Search $12-15 per 1k requests. Integrations include LangChain, CrewAI, OpenAI Tool Calling, and MCP server support (including Exa Agent and Connect). Compared to Perplexity or Brave, Exa's focus on agent-ready structured output and low latency makes it the preferred choice for production AI systems.
Behind the Verdict
Exa has carved a specific niche: web search that speaks the language of AI agents. Most search APIs return links and snippets; Exa returns structured data, highlights, and even full reasoning runs. That's a meaningful difference when your model needs to act on the results, not just display them. The 90% token reduction on highlights is the feature we'd reach for first—it directly cuts your LLM costs, which often exceed API fees. In practice, we'd pick Exa for coding agents (it powers Devin), lead enrichment (HubSpot uses it), and any agent that needs real-time facts. The sub-180ms Instant Search keeps agent loops snappy, while Deep Search handles the heavy lifting when you need citations and multi-step research. Where Exa bites: pricing is usage-based, and costs can scale if you're not careful with Deep Search or high-volume requests. The free tier ($20 credit upfront, $10 monthly) is enough for prototyping, but production workloads will hit the meter quickly. Also, the legacy /research endpoint and fields are deprecated as of April 2026—you must migrate to Deep Search, which is more powerful but pricier. Compared to alternatives: Perplexity's API is fine for consumer-grade answers but lacks Exa's structured outputs and latency guarantees. Brave's API is cheaper for simple web queries but doesn't offer the same agent-centric features. If you need raw HTML crawling at massive scale, you might pair Exa with a dedicated crawler, but for most agent use cases Exa is the complete package. We'd watch the index scale trajectory—Exa claims 1.4T URLs tracked and 80B documents served, on track to exceed Google-scale by 2027—so reliability should improve. Bottom line: Exa is the right choice if your AI agent's value depends on current, structured web data. It's not for static
Researching Exa? Get your full AI stack in 60 seconds.
Free, no signup — tell us your goal and get tools matched to your budget & existing stack.
Real-world workflow fit
Concrete scenarios for the personas Exa actually fits — and what changes day-one when you adopt it.
Integrating real-time search into a coding assistant to fetch current docs and API references.
Outcome: Agent can query Exa with sub-200ms latency, get token-efficient excerpts, and reduce context bloat while keeping code accurate.
Building a tool that turns 'AI startups in healthcare' into a structured list with funding data.
Outcome: Exa's structured outputs provide company names, CEOs, and funding details directly, eliminating manual parsing.
Setting up monitors to track competitor news and receive webhook alerts.
Outcome: Monitors run scheduled searches, deduplicate results, and push updates to Slack or other endpoints automatically.
Use Cases
- Power a coding agent with real-time search across docs, repos, and Stack Overflow.
- Build a sales-research tool that turns 'AI startups in healthcare' into a structured company list with CEO names and funding.
- Add neural retrieval to a RAG pipeline for higher quality grounding than keyword search.
- Create monitors that watch for new web content matching specific queries and send webhook alerts.
- Enrich CRM data with company and people intelligence using structured output schemas.
- Build a deep research assistant that synthesizes information from multiple web sources with citations.
- Automate news monitoring and competitive analysis with scheduled searches and summaries.
- Enable a voice agent to fetch up-to-date information in real time.
Limitations
- Pricing is per-call and can stack up on agentic workloads that issue many searches per task — set rate limits and cache aggressively.
- Neural search quality depends on phrasing; keyword fallback exists.
- Index coverage on very fresh news (last few hours) and non-English content is thinner than Google.
- Free tier provides $20 credits on sign-up and $10 credits per month.
- As of April 2026, the /research endpoint, resolvedSearchType, highlightScores, startCrawlDate, and endCrawlDate are deprecated and will sunset May 1, 2026.
- Deep search modes take up to 15 seconds per request, which may be too slow for real-time chat applications.
as of 2026-08-01
Verification history
We have re-verified Exa 15 times since . Each pass re-reads the vendor's own pages and re-checks every listed field against that evidence; passes where nothing had changed are marked as such.
- — re-verified summary, description, our verdict, our analysis, pricing model, pricing tiers, features, integrations, who it suits, who should skip it
- — re-verified summary, description, our verdict, our analysis, pricing model, pricing tiers, features, integrations, who it suits, who should skip it
- — re-verified summary, description, our verdict, our analysis, pricing model, pricing tiers, features, integrations, who it suits, who should skip it
- — re-verified summary, description, our verdict, our analysis, pricing model, pricing tiers, features, integrations, who it suits, who should skip it
- — re-verified summary, description, our verdict, our analysis, pricing model, pricing tiers, features, integrations, who it suits, who should skip it
- — re-verified summary, description, our verdict, our analysis, pricing model, pricing tiers, features, integrations, who it suits, who should skip it
Showing the 6 most recent of 15 verification passes.
Free to cite with attribution — this page re-verifies continuously.
12-month cost
Project the real annual outlay, including the implied monthly cost when only an annual tier is published.
Vendor list price only. Add-on usage, seat overages, and contract minimums are surfaced under Hidden costs & gotchas.
Plans compared
For each published Exa tier: who it actually fits, and what it adds vs. the previous tier. Cross-reference the cost calculator above for projected annual outlay.
Free
$0/mo
Ideal for
Developers exploring the API with $20 initial credits and $10 monthly, perfect for prototyping small projects.
What this tier adds
Free entry point with $20 credits on sign-up and $10 monthly credits; includes web search, highlights, and configurable latency up to 1s.
Search
$7/1k requests
Ideal for
Teams needing real-time web search with token-efficient page contents for high-volume requests.
What this tier adds
Paid tier at $7/1k requests; includes up to 10 results per request (additional results $1/1k) and same latency options as Free.
Agent
$0.012–$1.00/run
Ideal for
Developers building async research agents that require structured outputs and citations, with predictable per-run costs.
What this tier adds
Fixed effort modes from minimal ($0.012) to x-high ($1.00); includes structured outputs and access to data providers.
Contents
$1/1k pages per content type
Ideal for
Projects needing full page contents for LLM context, with token-efficient highlights and livecrawl policies.
What this tier adds
$1 per 1k pages per content type; includes full text and highlights, plus configurable livecrawl.
Deep Search
$12–15/1k requests
Ideal for
Complex research queries requiring multi-step agent workflows with grounded citations and structured outputs.
What this tier adds
$12–15 per 1k requests; includes multi-step reasoning, answers with structured outputs, and web-grounded citations.
Monitors
$15/1k requests
Ideal for
Organizations needing scheduled web monitoring for fresh events with webhook delivery and deduplication.
What this tier adds
$15 per 1k requests; includes scheduled searches, webhooks, and deduplication of recent events.
Enterprise
Custom
Ideal for
Large enterprises requiring custom rate limits, zero data retention, and dedicated support with volume discounts.
What this tier adds
Custom pricing; includes up to 1,000 results per search, custom index, SSO, and ZDR.
Where the pricing makes sense
The company stage and team size where Exa's pricing actually pencils out — and where peers do it cheaper.
Exa's pay-as-you-go pricing suits startups and developers who need scalable search API without upfront costs. For high-volume use, consider dedicated enterprise plans, which may offer volume discounts. Compared to Perplexity or Brave APIs, Exa's pricing is competitive for agentic workloads due to token-efficient highlights reducing LLM costs.
Setup time & first value
How long it actually takes to get something useful out of Exa — broken out by persona, not the marketing-page minute.
For developers, the Dashboard Onboarding generates a working integration snippet in under a minute. For basic search API calls, expect under 15 minutes including API key setup.
Switching to or from Exa
How to bring data in from common predecessors and how to get it back out — written for the switcher, not the buyer.
- →From Bing Search API: Exa's API is a drop-in replacement for web search endpoints; use the migration guide to adapt parameters and response formats.
- →From SerpAPI: Replace Google SERP scraping with Exa's structured search and highlights for LLM-friendly output.
- ↗To Perplexity API: If you need a hosted research assistant with built-in UI, consider Perplexity's higher-level API.
- ↗To Brave Search API: For low-cost general web search without structured outputs, Brave may be cheaper.
Integrations
Resources & Guides
Tutorials & Learning
Official links
Featured Head-to-Head Comparisons
Exa vs Firecrawl
For structured, token-efficient search with low latency and enterprise-grade features (SOC 2, SSO), Exa is the stronger choice, especially with its new Agent API and Deep Agent. For teams needing full web scraping, interaction, and an open-source, freemium starting point, Firecrawl offers unmatched flexibility with its keyless tier and broad integrations. Choose Exa for production AI agent search; choose Firecrawl for flexible data gathering and scraping at scale.
Exa vs Tavily
For AI agent builders needing the fastest real-time search with the broadest integration ecosystem and security filters, Tavily's freemium model and innovative x402 payments make it the more future-proof choice. However, if your priority is structured data extraction, lead enrichment, or optimizing LLM token costs, Exa's semantic search and Highlights provide a more specialized, production-ready solution. Choose based on your primary need: raw speed and integration breadth (Tavily) or semantic precision and token efficiency (Exa).
Popular in Web Scraping & Search APIs
Spider Cloud
AI web scraping API that turns any site into clean markdown or JSON for agents and RAG.
Mixpeek
Multimodal video search API: find any scene by description in your object storage.
Thunderbit
Agentic web scraper extension that turns any page into structured data in one click — no CSS selectors.
Frequently Asked Questions
Categories
Used Exa? Help shape our editorial sentiment research.


