4 papers
Facts Without Rules: Boundary Metadata Collapse in Multi-Agent LLM Handoffs
Yian Wang, Agam Goyal, Eshwar Chandrasekharan +1
Multi-agent LLM systems often coordinate by compressing an upstream interaction into a handoff artifact that downstream agents treat as shared state. We show that this handoff step…
Reasoning Consensus: Structural Ensembling of LLM Reasoning via Weighted DAG Aggregation
Amruta Parulekar, Jinu Lee, Dilek Hakkani-Tür +1
Large Language Models (LLMs) explore problems through chain-of-thought, but this exploration is buried in unstructured prose. On high-stakes tasks, users cannot tell which steps ar…
Masking or Mitigating? Deconstructing the Impact of Query Rewriting on Retriever Biases in RAG
Agam Goyal, Koyel Mukherjee, Apoorv Saxena +3
Dense retrievers in retrieval-augmented generation (RAG) systems exhibit systematic biases -- including brevity, position, literal matching, and repetition biases -- that can compr…
CausalDetox: Causal Head Selection and Intervention for Language Model Detoxification
Yian Wang, Yuen Chen, Agam Goyal +1
Large language models (LLMs) frequently generate toxic content, posing significant risks for safe deployment. Current mitigation strategies often degrade generation quality or requ…