collaborators

5 papers

cs.IR2026

MUSES: A Benchmark for Prospective Intellectual-Roots Retrieval

Rohan Pandey, Sunjae Kwon, Hong Yu

Scientific discovery depends on finding prior literature that shapes what comes next. Existing retrieval systems optimize for relevance and popularity, often favoring central paper…

cs.CL2026

Counterfactual Graph for Multi-Agent LLM Calibration

Jiatan Huang, Mingchen Li, Ziming Li +3

Multi-agent LLM systems often treat agreement as evidence: when many agents in a panel give the same answer, that answer is assumed to be more reliable. We show that this assumptio…

cs.CL2026

RICE-PO: Turning Retrieval Interactions into Credit Signals for Reasoning Agents

Mingchen Li, Hansi Zeng, Zhuo Qian +4

Retrieval is increasingly moving from one-shot matching toward interactive reasoning, where language agents iteratively inspect evidence, reformulate queries, and search again. Tra…

cs.CL2025

DischargeSim: A Simulation Benchmark for Educational Doctor-Patient Communication at Discharge

Zonghai Yao, Michael Sun, Won Seok Jang +3

Discharge communication is a critical yet underexplored component of patient care, where the goal shifts from diagnosis to education. While recent large language model (LLM) benchm…

cs.CL2025

Enhancing LLMs for Identifying and Prioritizing Important Medical Jargons from Electronic Health Record Notes Utilizing Data Augmentation: A Comparative Study

Won Seok Jang, Sharmin Sultana, Zonghai Yao +4

OpenNotes gives patients access to their EHR notes, but dense medical jargon limits comprehension. We evaluate closed-source and open-source LLMs for extracting and prioritizing th…