← Back to Model Beat
Models·2d ago·all news from July 19, 2026

Perplexity AI Releases WANDR: An Open Benchmark Evaluating Research Agents That Must Search Wide And Deep

Perplexity AI has released WANDR, an open-source evaluation benchmark designed to test the ability of research agents to gather evidence-based information across complex topics. The tool consists of 500 tasks that require agents to identify multiple entities and provide verifiable citations for each. By creating a standardized way to measure retrieval accuracy and depth, this benchmark aims to improve how developers evaluate the reliability of automated research systems.

Covered by 1 source

Related stories

ModelsApple Gets Approval for iPhone AI in China With Alibaba, BaiduJul 15 · 53 sourcesModelsChina Dismisses Claim that It Illicitly Extracts Foreign AI TechJul 17 · 86 sourcesModelsAlibaba’s Qwen Unveils Preview of Flagship AI ModelJul 19 · 38 sourcesModelsKimi's open model K3 nears GPT-5.6 Sol and Fable 5 while signaling the end of super cheap Chinese AIJul 16 · 212 sources