3 papers
cs.CY2026
Collaborative Disagreement Resolution for Scalable Oversight
Yuyang Jiang, Chacha Chen, Teng Wu +4
Debate, where AI agents argue opposing positions, has emerged as a key approach to scalable oversight. However, debate faces a fundamental tension: models are incentivized to be pe…
cs.HC2025
Beyond One-Way Influence: Bidirectional Opinion Dynamics in Multi-Turn Human-LLM Interactions
Yuyang Jiang, Longjie Guo, Yuchen Wu +3
Large language model (LLM)-powered chatbots are increasingly used for opinion exploration. Prior research examined how LLMs alter user views, yet little work extended beyond one-wa…
cs.CL2025
CLEAR: A Clinically-Grounded Tabular Framework for Radiology Report Evaluation
Yuyang Jiang, Chacha Chen, Shengyuan Wang +8
Existing metrics often lack the granularity and interpretability to capture nuanced clinical differences between candidate and ground-truth radiology reports, resulting in suboptim…