4 papers · 1 filter
VersaPRM: Multi-Domain Process Reward Model via Synthetic Reasoning Data
Thomas Zeng, Shuibai Zhang, Shutong Wu +13
Process Reward Models (PRMs) have proven effective at enhancing mathematical reasoning for Large Language Models (LLMs) by leveraging increased inference-time computation. However,…
Looped Transformers for Length Generalization
Ying Fan, Yilun Du, Kannan Ramchandran +1
Recent work has shown that Transformers trained from scratch can successfully solve various arithmetic and algorithmic tasks, such as adding numbers and computing parity. While the…
Learning to Understand: Identifying Interactions via the Möbius Transform
Justin S. Kang, Yigit E. Erginbas, Landon Butler +2
One of the key challenges in machine learning is to find interpretable representations of learned functions. The Möbius transform is essential for this purpose, as its coefficient…
The Fair Value of Data Under Heterogeneous Privacy Constraints in Federated Learning
Justin Kang, Ramtin Pedarsani, Kannan Ramchandran
Modern data aggregation often involves a platform collecting data from a network of users with various privacy options. Platforms must solve the problem of how to allocate incentiv…