6 papers
Fair on the Surface? Benchmarking Hidden-Output Fairness Gaps in LLM Recommenders
Chan Aristella Lu, Arya Fayyazi, Junhao Zhang +6
Fairness audits for LLM-based recommenders have largely focused on observable outputs, implicitly assuming that stable recommendations reflect stable internal processing. We challe…
FragFuse: Bypassing Access Control of Large Language Model Agents via Memory-Based Query Fragmentation and Fusion
Zixin Rao, Wentian Zhu, Chan Aristella Lu +5
Large language model (LLM) agents increasingly rely on long-term memory to support complex task execution, user personalization, and domain adaptation. Meanwhile, emerging access-c…
ShieldNet: Network-Level Guardrails against Emerging Supply-Chain Injections in Agentic Systems
Zhuowen Yuan, Zhaorun Chen, Zhen Xiang +5
Existing research on LLM agent security mainly focuses on prompt injection and unsafe input/output behaviors. However, as agents increasingly rely on third-party tools and MCP serv…
CDR-Agent: Intelligent Selection and Execution of Clinical Decision Rules Using Large Language Model Agents
Zhen Xiang, Aliyah R. Hsu, Austin V. Zane +6
Clinical decision-making is inherently complex and fast-paced, particularly in emergency departments (EDs) where critical, rapid and high-stakes decisions are made. Clinical Decisi…
Large Language Model Empowered Privacy-Protected Framework for PHI Annotation in Clinical Notes
Guanchen Wu, Linzhi Zheng, Han Xie +7
The de-identification of private information in medical data is a crucial process to mitigate the risk of confidentiality breaches, particularly when patient personal details are n…
X-Guard: Multilingual Guard Agent for Content Moderation
Bibek Upadhayay, Vahid Behzadan, Ph. D
Large Language Models (LLMs) have rapidly become integral to numerous applications in critical domains where reliability is paramount. Despite significant advances in safety framew…