2 papers
cs.LG2025
Data Mixture Optimization: A Multi-fidelity Multi-scale Bayesian Framework
Thomson Yen, Andrew Wei Tung Siah, Haozhe Chen +3
Careful curation of data sources can significantly improve the performance of LLM pre-training, but predominant approaches rely heavily on intuition or costly trial-and-error, maki…
cs.LG2024
PersonalLLM: Tailoring LLMs to Individual Preferences
Thomas P. Zollo, Andrew Wei Tung Siah, Naimeng Ye +2
As LLMs become capable of complex tasks, there is growing potential for personalized interactions tailored to the subtle and idiosyncratic preferences of the user. We present a pub…