2 papers
cs.LG2026
Pareto-Optimal Offline Reinforcement Learning via Smooth Tchebysheff Scalarization
Aadyot Bhatnagar, Peter Mørch Groth, Ali Madani
Large language models can be aligned with human preferences through offline reinforcement learning (RL) on small labeled datasets. While single-objective alignment is well-studied,…
q-bio.BM2025
Function-Guided Conditional Generation Using Protein Language Models with Adapters
Jason Yang, Aadyot Bhatnagar, Jeffrey A. Ruffolo +1
The conditional generation of proteins with desired functions is a key goal for generative models. Existing methods based on prompting of protein language models (PLMs) can generat…