Skip to main content

News

  • Perplexity
  • Q2D-Web
  • Agents

Perplexity introduces Q2D-Web to evaluate what research agents can retrieve

Perplexity's benchmark uses 190 million documents and nearly 70,000 agent-reformulated queries across ten languages. The leaderboard is public, but the benchmark is private; retrieval scores do not measure end-to-end answer accuracy.

Published (date only)Updated 3 sources

Primary source

Perplexity

Read the original: Q2D-Web announcement and evaluation methodology

Opens Perplexity in a new tab. Read it there before you rely on the summary above.

Unlock the full brief free.

  • What to check before you trust this story, written down
  • Each verified source, with why it matters and who published it
  • A note whenever a source has been withdrawn
  • Thirty days of stories to browse, not seven

This is not an account: there is no password, and the unlock is a cookie in this browser. You also join the Rise Productive newsletter from Demetri Panici, about once a week: what I built and what changed in AI. We'll email you a link to confirm, and you can unsubscribe in one click. The same signup unlocks every free tool on the site. How your email is handled.

Perplexity introduces Q2D-Web to evaluate what research agents can… | Rise Productive