3 papers
cs.LG2025
RL-Guided Data Selection for Language Model Finetuning
Animesh Jha, Harshit Gupta, Ananjan Nandi
Data selection for finetuning Large Language Models (LLMs) can be framed as a budget-constrained optimization problem: maximizing a model's downstream performance under a strict tr…
cs.LG2025
CTC-DRO: Robust Optimization for Reducing Language Disparities in Speech Recognition
Martijn Bartelds, Ananjan Nandi, Moussa Koulako Bala Doumbouya +3
Modern deep learning models often achieve high overall performance, but consistently fail on specific subgroups. Group distributionally robust optimization (group DRO) addresses th…
cs.CL2024
Sneaking Syntax into Transformer Language Models with Tree Regularization
Ananjan Nandi, Christopher D. Manning, Shikhar Murty
While compositional accounts of human language understanding are based on a hierarchical tree-like process, neural models like transformers lack a direct inductive bias for such tr…