2 papers
cs.AI2025
Parallel Key-Value Cache Fusion for Position Invariant RAG
Philhoon Oh, Jinwoo Shin, James Thorne
Recent advancements in Large Language Models (LLMs) underscore the necessity of Retrieval Augmented Generation (RAG) to leverage external information. However, LLMs are sensitive t…
cs.CL2024
Context Filtering with Reward Modeling in Question Answering
Sangryul Kim, James Thorne
Question Answering (QA) in NLP is the task of finding answers to a query within a relevant context retrieved by a retrieval system. Yet, the mix of relevant and irrelevant informat…