AI Matters

Why it matters

Perplexity releases open benchmark for deep-search AI agents

MarkTechPost Perplexity AI Releases WANDR: An Open Benchmark Evaluating Research Agents That Must Search Wide And Deep

Perplexity released WANDR, an open-source evaluation benchmark featuring 500 evidence-heavy tasks to test the search capabilities of AI agents.

Why it matters

It provides developers with a standardized way to test and improve AI agents that must perform complex, multi-step web research.