4 papers
WARP: Wasserstein-Aligned RAG for Population Opinions
Aman Singh Thakur, Aditya Agrawal, Alwarappan Nakkiran +1
RAG systems are increasingly used to summarize what large collections of documents say. A user asks "What do people think about X?" and receives an answer that reads as consensus.…
Training Leaves Traces: Centered Residual Signatures for Language Model Lineage Verification
Aman Singh Thakur, Rayan Khoury
Open-weight language models are fine-tuned, quantized, pruned, and merged, yet their provenance is often undocumented. We study data-free white-box lineage verification: can weight…
Retrieval-Augmented Generation Must Move Beyond Factual Grounding to Represent Diverse Opinions
Aditya Agrawal, Alwarappan Nakkiran, Darshan Fofadiya +3
This position paper argues that Retrieval-Augmented Generation (RAG) systems exhibit a factual bias-optimizing for epistemic uncertainty reduction while ignoring the aleatoric unce…
Judging the Judges: Evaluating Alignment and Vulnerabilities in LLMs-as-Judges
Aman Singh Thakur, Kartik Choudhary, Venkat Srinik Ramayapally +2
Offering a promising solution to the scalability challenges associated with human evaluation, the LLM-as-a-judge paradigm is rapidly gaining traction as an approach to evaluating l…