2 papers
cs.LG2025
Humanity's Last Exam
Long Phan, Alice Gatti, Ziwen Han +1144
Benchmarks are important tools for tracking the rapid advancements in large language model (LLM) capabilities. However, benchmarks are not keeping pace in difficulty: LLMs now achi…
math-ph2024
Infinitesimal and infinite numbers in applied mathematics
Aleksandr Bryzgalov, Kevin Islami, Paolo Giordano
The need to describe abrupt changes or response of nonlinear systems to impulsive stimuli is ubiquitous in applications. Also the informal use of infinitesimal and infinite quantit…