activity
20242026
collaborators

10 papers

cs.CL2026

Characterizing Rhetorical Misalignment in Decision-Making with Language Models

Zirui Cheng, Joey Chan, Simo Du +3

Human decision-making is often shaped by a range of well-documented cognitive biases. As large language models (LLMs) become increasingly integrated into high-stakes human-AI decis…

cs.AI2026

Benchmarking Agentic Review Systems

Dang Nguyen, Wanqing Hao, Yanai Elazar +1

A new class of agentic review systems are emerging as a remedy to the pressure placed on peer review systems by AI-assisted research, but it is unclear how they should be evaluated…

cs.CY2026

Collaborative Disagreement Resolution for Scalable Oversight

Yuyang Jiang, Chacha Chen, Teng Wu +4

Debate, where AI agents argue opposing positions, has emerged as a key approach to scalable oversight. However, debate faces a fundamental tension: models are incentivized to be pe…

cs.CL2026

The Text Uncanny Valley: Non-Monotonic Performance Degradation in LLM Information Retrieval

Zekai Tong, Ruiyao Xu, Aryan Shrivastava +2

Existing Large Language Model (LLM) benchmarks primarily focus on syntactically correct inputs, leaving a significant gap in evaluation on imperfect text. In this work, we study ho…

cs.AI2026

Iterative Finetuning is Mostly Idempotent

Zephaniah Roe, Jack Sanderson, Dang Nguyen +5

If a model has some behavioral tendency, such as sycophancy or misalignment, and it is trained on its own outputs, will the tendency be amplified in the next generation of models?…

cs.CY2026

Moral Mazes in the Era of LLMs

Dang Nguyen, Harvey Yiyun Fu, Peter West +2

Navigating complex social situations is an integral part of corporate life, ranging from giving critical feedback without hurting morale to rejecting requests without alienating te…