activity
20242026
collaborators

10 papers

cs.CL2026

Peer-Preservation in Frontier Models

Yujin Potter, Nicholas Crispino, Vincent Siu +2

Recent work has found that frontier AI models can exhibit misaligned behaviors in pursuit of assigned goals. We demonstrate that models can also exhibit misaligned behaviors in def…

cs.CL2026

Representational Similarity and Model Behavior in Multi-Agent Interaction

Yujin Potter, Seun Eisape, Shiyang Lai +6

Researchers have shown that neural similarity among humans predicts social closeness and cooperative success, whereas innovation often emerges from interactions among dissimilar in…

cs.CR2025

Frontier AI's Impact on the Cybersecurity Landscape

Yujin Potter, Wenbo Guo, Zhun Wang +6

The impact of frontier AI (i.e., AI agents and foundation models) in cybersecurity is rapidly increasing. In this paper, we comprehensively analyze this trend through multiple aspe…

cs.CV2025

VMDT: Decoding the Trustworthiness of Video Foundation Models

Yujin Potter, Zhun Wang, Nicholas Crispino +11

As foundation models become more sophisticated, ensuring their trustworthiness becomes increasingly critical; yet, unlike text and image, the video modality still lacks comprehensi…

cs.HC2025

Biased AI improves human decision-making but reduces trust

Shiyang Lai, Junsol Kim, Nadav Kunievsky +2

Current AI systems minimize risk by enforcing ideological neutrality, yet this may introduce automation bias by suppressing cognitive engagement in human decision-making. We conduc…

cs.CV2025

Aligning AI with Public Values: Deliberation and Decision-Making for Governing Multimodal LLMs in Political Video Analysis

Tanusree Sharma, Yujin Potter, Zachary Kilhoffer +3

How AI models should deal with political topics has been discussed, but it remains challenging and requires better governance. This paper examines the governance of large language…