23 citations · 24 across the 7 of their papers we have counts for
Showing 2025Show all
2 papers · 1 filter
cs.AI2025
Shutdownable Agents through POST-Agency
Elliott Thornley
Many fear that future artificial agents will resist shutdown. I present an idea - the POST-Agents Proposal - for ensuring that doesn't happen. I propose that we train agents to sat…
cs.LG2025★ 23 cited
Humanity's Last Exam
Long Phan, Alice Gatti, Ziwen Han +1144
Benchmarks are important tools for tracking the rapid advancements in large language model (LLM) capabilities. However, benchmarks are not keeping pace in difficulty: LLMs now achi…