3 papers
cs.LG2025
Data Mixture Optimization: A Multi-fidelity Multi-scale Bayesian Framework
Thomson Yen, Andrew Wei Tung Siah, Haozhe Chen +3
Careful curation of data sources can significantly improve the performance of LLM pre-training, but predominant approaches rely heavily on intuition or costly trial-and-error, maki…
cs.LG2025
PersonalLLM: Tailoring LLMs to Individual Preferences
Thomas P. Zollo, Andrew Wei Tung Siah, Naimeng Ye +2
As LLMs become capable of complex tasks, there is growing potential for personalized interactions tailored to the subtle and idiosyncratic preferences of the user. We present a pub…
stat.ML2024
Exchangeable Sequence Models Quantify Uncertainty Over Latent Concepts
Naimeng Ye, Hongseok Namkoong
Intelligent agents must be able to articulate its own uncertainty. In this work, we show that pre-trained sequence models are naturally capable of probabilistic reasoning over exch…