- Perplexity
- Q2D-Web
- Agents
Perplexity introduces Q2D-Web to evaluate what research agents can retrieve
Perplexity's benchmark uses 190 million documents and nearly 70,000 agent-reformulated queries across ten languages. The leaderboard is public, but the benchmark is private; retrieval scores do not measure end-to-end answer accuracy.
Published (date only)Updated 3 sources
Primary source
Perplexity
Read the original: Q2D-Web announcement and evaluation methodologyOpens Perplexity in a new tab. Read it there before you rely on the summary above.
Unlock the full brief free.
- What to check before you trust this story, written down
- Each verified source, with why it matters and who published it
- A note whenever a source has been withdrawn
- Thirty days of stories to browse, not seven
Related on Rise Productive
AI model picker
A free tool for choosing the AI setup that fits your work.
The newsletter
What I built and what changed in AI, about once a week.
