11 papers
Few shot chain-of-thought driven reasoning to prompt LLMs for open ended medical question answering
Saeel Sandeep Nachane, Ojas Gramopadhye, Prateek Chanda +5
In this paper, we propose a modified version of the MedQA-USMLE dataset, named MEDQA-OPEN, which contains open-ended medical questions without options to mimic clinical scenarios,…
Minibatch Selection for Language Models via Partition Matroid Constrained Gradient Matching
Prayas Agrawal, Prateek Chanda, Ishita Khatri +3
Training large language models (LLMs) on heterogeneous data requires selecting minibatches that balance convergence speed with coverage across domains. Existing methods either sele…
Online Distributional Prediction via Latent Cluster Geometry Under Drift and Corruption
Navyansh Mahla, Prateek Chanda, Ganesh Ramakrishnan
Online learning in non-stationary streams is often formulated as tracking a point estimate, but many applications require predicting the full data-generating distribution. We study…
Learning Task Mixtures from Task Affinities: A Probabilistic Graphical Model for Supervised Fine-Tuning
Prateek Chanda, Saral Sureka, Parth Pratim Chatterjee +3
Supervised fine-tuning performance for large language models depends strongly on how training budget is distributed across a heterogeneous set of tasks. In practice, mixtures are o…
UniPROT: Uniform Prototype Selection via Partial Optimal Transport with Submodular Guarantees
Prateek Chanda, Prayas Agrawal, Karthik S. Gurumoorthy +3
Selecting prototypical examples from a source distribution to represent a target data distribution is a fundamental problem in machine learning. Existing subset selection methods o…
Bandit Guided Submodular Curriculum for Adaptive Subset Selection
Prateek Chanda, Prayas Agrawal, Saral Sureka +3
Traditional curriculum learning proceeds from easy to hard samples, yet defining a reliable notion of difficulty remains elusive. Prior work has used submodular functions to induce…