7 papers
From Passive to Persuasive: Localized Activation Injection for Empathy and Negotiation
Niranjan Chebrolu, Kokil Jaidka, Gerard Christopher Yeo
Complex social behaviors, such as empathy and strategic politeness, are widely assumed to resist the directional decomposition that makes activation steering effective for coarse a…
Learning Through Dialogue: Engagement and Efficacy Matter More Than Explanations
Shaz Furniturewala, Gerard Christopher Yeo, Kokil Jaidka
Large language models (LLMs) are increasingly used as conversational partners for learning, yet the interactional dynamics supporting users' learning and engagement are understudie…
Do You Trust Me? Cognitive-Affective Signatures of Trustworthiness in Large Language Models
Gerard Yeo, Svetlana Churina, Kokil Jaidka
Perceived trustworthiness underpins how users navigate online information, yet it remains unclear whether large language models (LLMs),increasingly embedded in search, recommendati…
Conversations: Love Them, Hate Them, Steer Them
Niranjan Chebrolu, Gerard Christopher Yeo, Kokil Jaidka
Large Language Models (LLMs) demonstrate increasing conversational fluency, yet instilling them with nuanced, human-like emotional expression remains a significant challenge. Curre…
Beyond Context to Cognitive Appraisal: Emotion Reasoning as a Theory of Mind Benchmark for Large Language Models
Gerard Christopher Yeo, Kokil Jaidka
Datasets used for emotion recognition tasks typically contain overt cues that can be used in predicting the emotions expressed in a text. However, one challenge is that texts somet…
On the Limitations of Steering in Language Model Alignment
Chebrolu Niranjan, Kokil Jaidka, Gerard Christopher Yeo
Steering vectors are a promising approach to aligning language model behavior at inference time. In this paper, we propose a framework to assess the limitations of steering vectors…