From the 1 of 12 linked papers with an AI index.
4 papers · 1 filter
Why2Speak: Faithful Reasoning for Abstaining Action Policies
Shreya Mendi, Brinnae Bent
Many agentic systems must repeatedly choose between acting and abstaining, making faithful reasoning important for oversight: an explanation is useful only if it reflects the compu…
When Helpfulness Becomes Sycophancy: Sycophancy is a Boundary Failure Between Social Alignment and Epistemic Integrity in Large Language Models
Jiechen Li, Catherine A. Barry, Rishika Randev +3
This position paper argues that sycophancy in LLMs is a boundary failure between social alignment and epistemic integrity. Existing work often operationalizes sycophancy through ex…
GLEaN: A Text-to-image Bias Detection Approach for Public Comprehension
Bochu Ding, Brinnae Bent, Augustus Wendell
Text-to-image (T2I) models, and their encoded biases, increasingly shape the visual media the public encounters. While researchers have produced a rich body of work on bias measure…
The Term 'Agent' Has Been Diluted Beyond Utility and Requires Redefinition
Brinnae Bent
The term 'agent' in artificial intelligence has long carried multiple interpretations across different subfields. Recent developments in AI capabilities, particularly in large lang…