Showing cs.CLShow all
2 papers · 1 filter
cs.CL2026
PQR: A Framework to Generate Diverse and Realistic User Queries that Elicit QA Agent Failures
Yunan Lu, Luigi Liu, Omar Yahia +2
Evaluating LLM-based agents remains challenging because identifying meaningful failure cases often requires substantial human effort to design realistic test scenarios. Prior works…
cs.CL2024
Alexpaca: Learning Factual Clarification Question Generation Without Examples
Matthew Toles, Yukun Huang, Zhou Yu +1
Real-life tasks such as giving legal or technical advice often lack complete context at the outset and can have disparate answers depending thereon. The ability to derive missing f…