2 papers
cs.LG2026
Humanity's Last Exam
Long Phan, Alice Gatti, Ziwen Han +1144
Benchmarks are important tools for tracking the rapid advancements in large language model (LLM) capabilities. However, benchmarks are not keeping pace in difficulty: LLMs now achi…
cs.GT2026
How to Tamper with a Parliament: Strategic Campaigns in Apportionment Elections
Robert Bredereck, Piotr Faliszewski, MichaÅ Furdyna +6
In parliamentary elections, parties compete for a limited, typically fixed number of seats. Most parliaments are assembled using apportionment methods that distribute the seats bas…