4 papers
Limitations of refinement methods for weak to strong generalization
Seamus Somerstep, Ya'acov Ritov, Mikhail Yurochkin +2
Standard techniques for aligning large language models (LLMs) utilize human-produced data, which could limit the capability of any aligned LLM to human level. Label refinement and…
CARROT: A Cost Aware Rate Optimal Router
Seamus Somerstep, Felipe Maia Polo, Allysson Flavio Melo de Oliveira +5
With the rapid growth in the number of Large Language Models (LLMs), there has been a recent interest in LLM routing, or directing queries to the cheapest LLM that can deliver a su…
Sloth: scaling laws for LLM skills to predict multi-benchmark performance across families
Felipe Maia Polo, Seamus Somerstep, Leshem Choshen +2
Scaling laws for large language models (LLMs) predict model performance based on parameters like size and training data. However, differences in training configurations and data pr…
The Supersingularity of Hurwitz Curves
Dean Bisogno, Erin Dawson, Henry Frauenhoff +5
We study when Hurwitz curves are supersingular. Specifically, we show that the curve , with and relatively prime, is s…