4 papers · 1 filter
Collaborative Disagreement Resolution for Scalable Oversight
Yuyang Jiang, Chacha Chen, Teng Wu +4
Debate, where AI agents argue opposing positions, has emerged as a key approach to scalable oversight. However, debate faces a fundamental tension: models are incentivized to be pe…
Moral Mazes in the Era of LLMs
Dang Nguyen, Harvey Yiyun Fu, Peter West +2
Navigating complex social situations is an integral part of corporate life, ranging from giving critical feedback without hurting morale to rejecting requests without alienating te…
On the Effectiveness and Generalization of Race Representations for Debiasing High-Stakes Decisions
Dang Nguyen, Chenhao Tan
Understanding and mitigating biases is critical for the adoption of large language models (LLMs) in high-stakes decision-making. We introduce Admissions and Hiring, decision tasks…
GPT-4V Cannot Generate Radiology Reports Yet
Yuyang Jiang, Chacha Chen, Dang Nguyen +2
GPT-4V's purported strong multimodal abilities raise interests in using it to automate radiology report writing, but there lacks thorough evaluations. In this work, we perform a sy…