2 papers
cs.CL2026
Transforming LLMs into Efficient Cross-Encoders via Knowledge Distillation for RAG Reranking
Shreeya Dasa Lakshminath, Shubhan S
Cross-encoders achieve high reranking accuracy in Retrieval-Augmented Generation (RAG) pipelines but impose quadratic inference costs that limit real-time deployment. We address th…
cs.LG2026
Nonparametric Bayesian Inverse Reinforcement Learning with Data-Parallel Gibbs Sampling
Sai Anirudh Katupilla, Shreeya Dasa Lakshminath
Inverse Reinforcement Learning recovers reward functions from expert demonstrations, but standard formulations assume that all demonstrations come from a single expert. When demons…