AI News
Perplexity's WANDR benchmark exposes a brutal truth: research agents can't yet assemble 170,000 evidence-backed records
MarkTechPost · Jul 19, 2026 · 3 min read
Most AI benchmarks are multiple-choice tests. They're tidy, they're narrow, and they don't resemble the sprawling infor...