4 papers
Do You Trust Me? Cognitive-Affective Signatures of Trustworthiness in Large Language Models
Gerard Yeo, Svetlana Churina, Kokil Jaidka
Perceived trustworthiness underpins how users navigate online information, yet it remains unclear whether large language models (LLMs),increasingly embedded in search, recommendati…
Beyond Context to Cognitive Appraisal: Emotion Reasoning as a Theory of Mind Benchmark for Large Language Models
Gerard Christopher Yeo, Kokil Jaidka
Datasets used for emotion recognition tasks typically contain overt cues that can be used in predicting the emotions expressed in a text. However, one challenge is that texts somet…
On the Limitations of Steering in Language Model Alignment
Chebrolu Niranjan, Kokil Jaidka, Gerard Christopher Yeo
Steering vectors are a promising approach to aligning language model behavior at inference time. In this paper, we propose a framework to assess the limitations of steering vectors…
Conversations: Love Them, Hate Them, Steer Them
Niranjan Chebrolu, Gerard Christopher Yeo, Kokil Jaidka
Large Language Models (LLMs) demonstrate increasing conversational fluency, yet instilling them with nuanced, human-like emotional expression remains a significant challenge. Curre…