3 papers
cs.RO2026
Evaluating Generative Models as Interactive Emergent Representations of Human-Like Collaborative Behavior
Shinas Shaji, Teena Chakkalayil Hassan, Sebastian Houben +1
Human-AI collaboration requires AI agents to understand human behavior for effective coordination. While advances in foundation models show promising capabilities in understanding…
cs.RO2025
Reliable Robotic Task Execution in the Face of Anomalies
Bharath Santhanam, Alex Mitrevski, Santosh Thoduka +2
Learned robot policies have consistently been shown to be versatile, but they typically have no built-in mechanism for handling the complexity of open environments, making them pro…
cs.CL2025
GG-BBQ: German Gender Bias Benchmark for Question Answering
Shalaka Satheesh, Katrin Klug, Katharina Beckh +3
Within the context of Natural Language Processing (NLP), fairness evaluation is often associated with the assessment of bias and reduction of associated harm. In this regard, the e…