2 papers
cs.AI2026
Learning When Not to Decide: A Framework for Overcoming Factual Presumptuousness in AI Adjudication
Mohamed Afane, Emily Robitschek, Derek Ouyang +1
A well-known limitation of AI systems is presumptuousness: the tendency of AI systems to provide confident answers when information may be lacking. This challenge is particularly a…
cs.CL2026
Benchmarking Legal RAG: The Promise and Limits of AI Statutory Surveys
Mohamed Afane, Emaan Hariri, Derek Ouyang +1
Retrieval-augmented generation (RAG) offers significant potential for legal AI, yet systematic benchmarks are sparse. Prior work introduced LaborBench to benchmark RAG models based…