2 papers
cs.LG2026
DRACO: a Cross-Domain Benchmark for Deep Research Accuracy, Completeness, and Objectivity
Joey Zhong, Hao Zhang, Clare Southern +7
We present DRACO (Deep Research Accuracy, Completeness, and Objectivity), a benchmark of complex deep research tasks. These tasks, which span 10 domains and draw on information sou…
cs.LG2025
The Adoption and Usage of AI Agents: Early Evidence from Perplexity
Jeremy Yang, Noah Yonack, Kate Zyskowski +3
This paper presents the first large-scale field study of the adoption, usage intensity, and use cases of general-purpose AI agents operating in open-world web environments. Our ana…