Perplexity's benchmark and dataset for evaluating deep, wide research work, comprising 500 tasks represented as hierarchical, independently verifiable source-backed records.

Recent stories
1 linked story
Perplexity's benchmark and dataset for evaluating deep, wide research work, comprising 500 tasks represented as hierarchical, independently verifiable source-backed records.
