# Exa launches ATLAS benchmark for agentic web search

> The new ATLAS benchmark grades how well AI agents retrieve complete, accurate results from real web searches, highlighting current cost and performance gaps.

Oossa · 2026-10-08 · https://oossa.com/en/exa-launches-atlas-benchmark-for-agentic-web-search

Exa AI announced ATLAS, a benchmark that measures the accuracy and completeness of AI agents that rely on web search. It uses 547 real‑world research tasks drawn from anonymized search demand. The benchmark shows that even the most expensive agents miss about a third of the correct results and that no system costing under $1 per task reaches a row F1 score above 0.5.

## The facts

- 547 tasks covering deep and wide research scenarios
- Agents costing less than $1 per task never exceed 0.5 row F1

## Why it matters

Developers can use ATLAS to identify which search backends and agent configurations give the best return for money, guiding more reliable AI‑driven research tools.

## Sources & references

1. [Introducing ATLAS: A new benchmark for agentic search](https://exa.ai/blog/atlas-benchmark) – Exa

Last updated: 2026-10-08
