8 citations · 14 across the 13 of their papers we have counts for
Showing cs.AIShow all
2 papers · 1 filter
cs.AI2026
SWE-chat: Coding Agent Interactions From Real Users in the Wild
Joachim Baumann, Vishakh Padmakumar, Xiang Li +3
AI coding agents are being adopted at scale, yet we lack empirical evidence on how people actually use them and how much of their output is useful in practice. We present SWE-chat,…
cs.AI2023★ 8 cited
Debate Helps Supervise Unreliable Experts
Julian Michael, Salsabila Mahdi, David Rein +4
As AI systems are used to answer more difficult questions and potentially help create new knowledge, judging the truthfulness of their outputs becomes more difficult and more impor…