2 papers
cs.AI2026
AttuneBench: A Conversation-Based Benchmark for LLM Emotional Intelligence
Kate M. Lubrano, Faisal Sayed, Ankita Rathod +4
Emotional intelligence (EI), the ability to perceive, understand, and respond appropriately to others' emotional states, is central to human communication, and increasingly importa…
cs.LG2026
LEAP: Trajectory-Level Evaluation of LLMs in Iterative Scientific Design
Marilyn Zhang, Tianfeng Chen, Fabián Barzuna +2
LLMs are increasingly deployed in autonomous laboratories, under the assumption that their domain priors and reasoning over iterative feedback let them converge on good designs in…