3 papers
cs.AI2026
ART: Action-based Reasoning Task Benchmarking for Medical AI Agents
Ananya Mantravadi, Shivali Dalmia, Abhishek Mukherji
Reliable clinical decision support requires medical AI agents capable of safe, multi-step reasoning over structured electronic health records (EHRs). While large language models (L…
cs.AI2025
LegalWiz: A Multi-Agent Generation Framework for Contradiction Detection in Legal Documents
Ananya Mantravadi, Shivali Dalmia, Olga Pospelova +3
Retrieval-Augmented Generation (RAG) integrates large language models (LLMs) with external sources, but unresolved contradictions in retrieved evidence often lead to hallucinations…
cs.CY2025
Scaling Success: A Systematic Review of Peer Grading Strategies for Accuracy, Efficiency, and Learning in Contemporary Education
Uchswas Paul, Ananya Mantravadi, Jash Shah +4
Peer grading has emerged as a scalable solution for assessment in large and online classrooms, offering both logistical efficiency and pedagogical value. However, designing effecti…