Showing cs.LGShow all
3 papers · 1 filter
cs.LG2026
Shallower ReLU Network Representations via Exact Linear Algebra
Kilian RueÃ, Gennadiy Averkov, Florestan Brunck +7
We prove that the maximum of real numbers is exactly representable by a ReLU network with two hidden layers for every . The constructions are obtained by reducing the…
cs.LG2026★ 18 cited
Humanity's Last Exam
Long Phan, Alice Gatti, Ziwen Han +1144
Benchmarks are important tools for tracking the rapid advancements in large language model (LLM) capabilities. However, benchmarks are not keeping pace in difficulty: LLMs now achi…
cs.LG2026
Better Neural Network Expressivity: Subdividing the Simplex
Egor Bakaev, Florestan Brunck, Christoph Hertrich +2
This work studies the expressivity of ReLU neural networks with a focus on their depth. A sequence of previous works showed that hidden layers are suffi…