3 papers
cs.AI2026
DI-Bench: Systematically Generating In-Domain Data Intelligence Benchmarks for Enterprise Agents
Jiangyun Zhang, Kristen Surrao, Torpong Nitayanont +10
Evaluating enterprise agents on domain-specific benchmarks is critical, yet public benchmarks rarely evaluate whether agents can integrate business knowledge with analytical comput…
cs.AI2026
DataSpace: Benchmarking Data Agents for Verifiable Analytics over Heterogeneous Workspaces
Boyan Li, Zhuowen Liang, Yupeng Xie +11
Data agents enable natural-language analytics over organizational workspaces, where relevant evidence may be scattered across databases, structured files, long documents, and multi…
cs.DB2025
AgenticData: An Agentic Data Analytics System for Heterogeneous Data
Ji Sun, Guoliang Li, Peiyao Zhou +3
Existing unstructured data analytics systems rely on experts to write code and manage complex analysis workflows, making them both expensive and time-consuming. To address these ch…