3 papers
cs.AI2026
RecourseBench: A Modular Framework for Reproducible Algorithmic Recourse Evaluation
Zahra Khotanlou, Hashir Ahmed, Chenghao Tan +2
Algorithmic recourse methods provide counterfactual explanations that inform individuals of the actions required to overturn an unfavorable model decision. Despite rapid methodolog…
cs.LG2026
A Unified Perturbation Framework for Analyzing Leaderboard Stability and Manipulation
Hosna Oyarhoseini, Jimmy Lin, Amir-Hossein Karimi
Evaluation leaderboards such as LMArena play a central role in benchmarking large language models by aggregating pairwise human preferences into model rankings, yet the robustness…
cs.MA2025
Robust Coordination under Misaligned Communication via Power Regularization
Nancirose Piazza, Amirhossein Karimia, Behnia Soleymanib +2
Effective communication in Multi-Agent Reinforcement Learning (MARL) can significantly enhance coordination and collaborative performance in complex and partially observable enviro…