2 papers
cs.CL2026
AdaMame: A Training Recipe for Adaptive Multilingual Reasoning
Dayeon Ki, Kevin Duh, Marine Carpuat
While Large Reasoning Models (LRMs) show strong performance in English, they often fail to reason in the language of the query, a phenomenon known as language collapse. Existing RL…
cs.CL2026
Data Kernel Perspective Space Performance Guarantees for Synthetic Data from Transformer Models
Michael Browder, Kevin Duh, J. David Harris +5
Scarcity of labeled training data remains the long pole in the tent for building performant language technology and generative AI models. Transformer models -- particularly LLMs --…