2 papers
cs.CL2025
Agent Bain vs. Agent McKinsey: A New Text-to-SQL Benchmark for the Business Domain
Yue Li, Ran Tao, Derek Hommel +4
Text-to-SQL benchmarks have traditionally only tested simple data access as a translation task of natural language to SQL queries. But in reality, users tend to ask diverse questio…
cs.CL2025
Show or Tell? Modeling the evolution of request-making in Human-LLM conversations
Shengqi Zhu, Jeffrey M. Rzeszotarski, David Mimno
Designing user-centered LLM systems requires understanding how people use them, but patterns of user behavior are often masked by the variability of queries. In this work, we intro…