3 papers
cs.AI2026
When Web Agents Finish but Still Fail: Reproducible Triggers and Trace Diagnostics for Parallel Web Exploration
Aagam Sogani, Botao Rui, Swetha Vaidyanathan +3
Long-horizon web agents often fail in ways hidden by final-answer evaluation: they may visit useful pages, produce a well-formed answer, and terminate confidently while still missi…
cs.IR2025
REaR: Retrieve, Expand and Refine for Effective Multitable Retrieval
Rishita Agarwal, Himanshu Singhal, Peter Baile Chen +3
Answering natural language queries over relational data often requires retrieving and reasoning over multiple tables, yet most retrievers optimize only for query-table relevance an…
cs.CY2025
REDDIX-NET: A Novel Dataset and Benchmark for Moderating Online Explicit Services
MSVPJ Sathvik, Manan Roy Choudhury, Rishita Agarwal +2
The rise of online platforms has enabled covert illicit activities, including online prostitution, to pose challenges for detection and regulation. In this study, we introduce REDD…