3 papers
cs.LG2025
ESLM: Risk-Averse Selective Language Modeling for Efficient Pretraining
Melis Ilayda Bal, Volkan Cevher, Michael Muehlebach
Large language model pretraining is compute-intensive, yet many tokens contribute marginally to learning, resulting in inefficiency. We introduce Efficient Selective Language Model…
cs.LG2025
Adversarial Training for Defense Against Label Poisoning Attacks
Melis Ilayda Bal, Volkan Cevher, Michael Muehlebach
As machine learning models grow in complexity and increasingly rely on publicly sourced data, such as the human-annotated labels used in training large language models, they become…
cs.LG2024
Optimistic Games for Combinatorial Bayesian Optimization with Application to Protein Design
Melis Ilayda Bal, Pier Giuseppe Sessa, Mojmir Mutny +1
Bayesian optimization (BO) is a powerful framework to optimize black-box expensive-to-evaluate functions via sequential interactions. In several important problems (e.g. drug disco…