4 papers
GrepSeek: Training Search Agents for Direct Corpus Interaction
Alireza Salemi, Chang Zeng, Atharva Nijasure +4
Large Language Model (LLM) search agents have shown strong promise for knowledge-intensive language tasks through multiple rounds of reasoning and information retrieval. Most exist…
Hedonic Neurons: A Mechanistic Mapping of Latent Coalitions in Transformer MLPs
Tanya Chowdhury, Atharva Nijasure, Yair Zick +1
Fine-tuned Large Language Models (LLMs) encode rich task-specific features, but the form of these representations, especially within MLP layers, remains unclear. Empirical inspecti…
How Relevance Emerges: Interpreting LoRA Fine-Tuning in Reranking LLMs
Atharva Nijasure, Tanya Chowdhury, James Allan
We conduct a behavioral exploration of LoRA fine-tuned LLMs for Passage Reranking to understand how relevance signals are learned and deployed by Large Language Models. By fine-tun…
Probing Ranking LLMs: A Mechanistic Analysis for Information Retrieval
Tanya Chowdhury, Atharva Nijasure, James Allan
Transformer networks, particularly those achieving performance comparable to GPT models, are well known for their robust feature extraction abilities. However, the nature of these…