2 papers
cs.LG2025
HARDMath2: A Benchmark for Applied Mathematics Built by Students as Part of a Graduate Class
James V. Roggeveen, Erik Y. Wang, Will Flintoft +42
Large language models (LLMs) have shown remarkable progress in mathematical problem-solving, but evaluation has largely focused on problems that have exact analytical solutions or…
cs.CL2024
LayerSkip: Enabling Early Exit Inference and Self-Speculative Decoding
Mostafa Elhoushi, Akshat Shrivastava, Diana Liskovich +10
We present LayerSkip, an end-to-end solution to speed-up inference of large language models (LLMs). First, during training we apply layer dropout, with low dropout rates for earlie…