4 papers
How Do Large Language Models Judge Social Attraction? Evidence from Theory-Grounded Persona Ratings Across Multiple LLMs and Humans
Hasan Mahmud, Khawaja Abaid Ullah, Mohammad Javad Khojasteh +2
Large language models (LLMs) are increasingly used to perform subjective evaluations traditionally made by humans, yet their validity as social judges remains unclear. This paper e…
When Robots Rate Their Own Interactions: Engagement Validity and the Strangeness Failure
Victor Lockwood, Hasan Mahmud, Mohammad Javad Khojasteh +2
Human-robot interaction (HRI) evaluation relies almost exclusively on human-completed questionnaires, leaving the robot's perspective unexamined. We propose an \textit{inverted eva…
Optimizing Neurorobot Policy under Limited Demonstration Data through Preference Regret
Viet Dung Nguyen, Yuhang Song, Anh Nguyen +3
Robot reinforcement learning from demonstrations (RLfD) assumes that expert data is abundant; this is usually unrealistic in the real world given data scarcity as well as high coll…
Robust Speech-Workload Estimation for Intelligent Human-Robot Systems
Julian Fortune, Julie A. Adams, Jamison Heard
Demanding task environments (e.g., supervising a remotely piloted aircraft) require performing tasks quickly and accurately; however, periods of low and high operator workload can…