3 papers
cs.IR2025
REaR: Retrieve, Expand and Refine for Effective Multitable Retrieval
Rishita Agarwal, Himanshu Singhal, Peter Baile Chen +3
Answering natural language queries over relational data often requires retrieving and reasoning over multiple tables, yet most retrievers optimize only for query-table relevance an…
cs.AI2025
Better Call CLAUSE: A Discrepancy Benchmark for Auditing LLMs Legal Reasoning Capabilities
Manan Roy Choudhury, Adithya Chandramouli, Mannan Anand +1
The rapid integration of large language models (LLMs) into high-stakes legal work has exposed a critical gap: no benchmark exists to systematically stress-test their reliability ag…
cs.CY2025
REDDIX-NET: A Novel Dataset and Benchmark for Moderating Online Explicit Services
MSVPJ Sathvik, Manan Roy Choudhury, Rishita Agarwal +2
The rise of online platforms has enabled covert illicit activities, including online prostitution, to pose challenges for detection and regulation. In this study, we introduce REDD…