works on

From the 1 of 11 linked papers with an AI index.

collaborators

11 papers

cs.CL2026

Value Drifts: Tracing Value Alignment During LLM Post-Training

Mehar Bhatia, Shravan Nayak, Gaurav Kamath +4

The paper studies how large language models acquire and change their alignment with human values during post‑training, analyzing the impact of supervised fine‑tuning and preference…

cs.MM2026

DetectZoo: A Unified Toolkit for AI-Generated Content Detection Across Text, Audio, and Image Modalities

Sajad Ebrahimi, Nima Jamali, Bardia Shirsalimian +8

The growing popularity and capacity of generative models have eroded the distinction between human and machine-generated content, motivating a growing body of work on detection acr…

cs.CL2026

Infusing Theory of Mind into Socially Intelligent LLM Agents

EunJeong Hwang, Yuwei Yin, Giuseppe Carenini +2

Theory of Mind (ToM)-an understanding of the mental states of others-is a key aspect of human social intelligence, yet, chatbots and LLM-based social agents do not typically integr…

cs.HC2026

Games That Teach, Chats That Convince: Comparing Interactive and Static Formats for Persuasive Learning

Seyed Hossein Alavi, Zining Wang, Shruthi Chockkalingam +2

Interactive systems such as chatbots and games are increasingly used to persuade and educate on sustainability-related topics, yet it remains unclear how different delivery formats…

cs.AI2026

Agents of Chaos

Natalie Shapira, Chris Wendler, Avery Yen +35

We report an exploratory red-teaming study of autonomous language-model-powered agents deployed in a live laboratory environment with persistent memory, email accounts, Discord acc…

cs.CL2025

BottleHumor: Self-Informed Humor Explanation using the Information Bottleneck Principle

EunJeong Hwang, Peter West, Vered Shwartz

Humor is prevalent in online communications and it often relies on more than one modality (e.g., cartoons and memes). Interpreting humor in multimodal settings requires drawing on…