3 papers
cs.AI2026
SafetyALFRED: Evaluating Safety-Conscious Planning of Multimodal Large Language Models
Josue Torres-Fonseca, Naihao Deng, Yinpei Dai +5
Multimodal Large Language Models are increasingly adopted as autonomous agents in interactive environments, yet their ability to proactively address safety hazards remains insuffic…
cs.LG2026
Identifying and Mitigating Gender Cues in Academic Recommendation Letters: An Interpretability Case Study
Charlotte S. Alexander, Shane Storks, Souradip Pal +4
Letters of recommendation (LoRs) can carry patterns of implicitly gendered language that can inadvertently influence downstream decisions, e.g. in hiring and admissions. In this wo…
cs.CL2025
Discovering Properties of Inflectional Morphology in Neural Emergent Communication
Miles Gilberti, Shane Storks, Huteng Dai
Emergent communication (EmCom) with deep neural network-based agents promises to yield insights into the nature of human language, but remains focused primarily on a few subfield-s…