← Back to Model Beat
Models·Jul 19·all news from July 19, 2026

Perplexity AI Releases WANDR: An Open Benchmark Evaluating Research Agents That Must Search Wide And Deep

Perplexity AI has released WANDR, an open-source evaluation benchmark designed to test the ability of research agents to gather evidence-based information across complex topics. The tool consists of 500 tasks that require agents to identify multiple entities and provide verifiable citations for each. By creating a standardized way to measure retrieval accuracy and depth, this benchmark aims to improve how developers evaluate the reliability of automated research systems.

Covered by 1 source

Related stories

ModelsMicrosoft and Mistral strike multi-billion-dollar deal to build AI infrastructure across EuropeJul 21 · 97 sourcesModelsAlibaba’s Qwen Unveils Preview of Flagship AI ModelJul 19 · 47 sourcesModelsOne tampered ChatGPT link could spawn a rogue AI agent that took orders from an attacker every five minutesJul 22 · 34 sourcesModelsChina Dismisses Claim that It Illicitly Extracts Foreign AI TechJul 17 · 94 sources