73 citations · 157 across the 37 of their papers we have counts for
17 papers · 1 filter
Benchmarking Hybrid Deep Research Across Database Querying and Web Search
Ruofan Wu, Peiran Xu, Xiaolong Li +9
While autonomous agents have made significant strides in "deep research" by iteratively navigating the open web to synthesize information, real-world problem-solving is rarely conf…
Spider 2.0-AIFunc: Extending Real-World Text-to-SQL to AI-Native SQL Workflows
Tianyang Liu, Canwen Xu, Fangyu Lei +6
Major cloud data platforms now expose large language model capabilities as native SQL functions, enabling analysts to perform classification, filtering, sentiment analysis, extract…
Learning to Retrieve: Dual-Level Long-Term Memory for Text-to-SQL Agents
Yibo Wang, Nikki Lijing Kuang, Philip S. Yu +2
Interactive text-to-SQL agents solve database tasks through multi-turn interactions involving schema exploration, query execution, feedback interpretation, and decision revision. L…
Residual Skill Optimization for Text-to-SQL Ensembles
Jiongli Zhu, Haoquan Guan, Parjanya Prajakta Prashant +8
Text-to-SQL ensembles improve over single-candidate generation by drawing multiple SQL candidates and selecting one, but their effectiveness is bounded by Pass@K, the probability t…
Learning to Self-Evolve
Xiaoyin Chen, Canwen Xu, Yite Wang +3
We introduce Learning to Self-Evolve (LSE), a reinforcement learning framework that trains large language models (LLMs) to improve their own contexts at test time. We situate LSE i…
MedOrch: Medical Diagnosis with Tool-Augmented Reasoning Agents for Flexible Extensibility
Yexiao He, Ang Li, Boyi Liu +2
Healthcare decision-making represents one of the most challenging domains for Artificial Intelligence (AI), requiring the integration of diverse knowledge sources, complex reasonin…