Settings

Theme

Wandr Benchmark: Evaluating Research Agents That Must Search Wide and Deep

research.perplexity.ai

1 points by tagawa · 1 comment

Reader

1 thread
tagawaOP

Repo with benchmark tasks, evaluation harness, tech report:

https://github.com/perplexityai/wandr

Keyboard Shortcuts

j
Next item
k
Previous item
o / Enter
Open selected item
?
Show this help
Esc
Close modal / clear selection