3 papers
cs.CL2025
Latent Traits and Cross-Task Transfer: Deconstructing Dataset Interactions in LLM Fine-tuning
Shambhavi Krishna, Atharva Naik, Chaitali Agarwal +3
Large language models are increasingly deployed across diverse applications. This often includes tasks LLMs have not encountered during training. This implies that enumerating and…
cs.LG2024
Solving the Inverse Alignment Problem for Efficient RLHF
Shambhavi Krishna, Aishwarya Sahoo
Collecting high-quality preference datasets for reinforcement learning from human feedback (RLHF) is resource-intensive and challenging. As a result, researchers often train reward…
cs.AI2024
PAFFA: Premeditated Actions For Fast Agents
Shambhavi Krishna, Zheng Chen, Yuan Ling +4
Modern AI assistants have made significant progress in natural language understanding and tool-use, with emerging efforts to interact with Web interfaces. However, current approach…