1 citations · 2 across the 3 of their papers we have counts for
3 papers
cs.SE2026
Anchored Self-Play for Code Repair
Caroline Choi, Zeyneb Kaya, Shirley Wu +3
Code repair is an important capability for language models (LMs): given a buggy program and unit tests, an LM must produce a fixed program that passes the tests. Because code repai…
cs.LG2025★ 1 cited
OpenThoughts: Data Recipes for Reasoning Models
Etash Guha, Ryan Marten, Sedrick Keh +47
Reasoning models have made rapid progress on many benchmarks involving math, code, and science. Yet, there are still many open questions about the best training recipes for reasoni…
cs.LG2024★ 1 cited
AutoFT: Learning an Objective for Robust Fine-Tuning
Caroline Choi, Yoonho Lee, Annie Chen +3
Foundation models encode rich representations that can be adapted to downstream tasks by fine-tuning. However, fine-tuning a model on one data distribution often degrades performan…