§ feed · storyline
WANDR Benchmark: Evaluating Research Agents That Must Search Wide and Deep
Perplexity releases WANDR benchmark for evaluating research agents' wide and deep search performance.
Perplexity released findings from the WANDR benchmark, which evaluates the performance of various research agents in wide and deep searching.
§ sources1 publication · timeline below
- research.perplexity.aiWANDR Benchmark: Evaluating Research Agents That Must Search Wide and Deepprimary