4 papers
When Does Multi-Agent Collaboration Help? An Entropy Perspective
Yuxuan Zhao, Sijia Chen, Ningxin Su
Multi-agent systems (MAS) have emerged as a prominent paradigm for leveraging large language models (LLMs) to tackle complex tasks. However, the mechanisms governing the effectiven…
LegalCiteBench: Evaluating Citation Reliability in Legal Language Models
Sijia Chen, Hang Yin, Shunfan Zhou
Large language models (LLMs) are increasingly integrated into legal drafting and research workflows, where incorrect citations or fabricated precedents can cause serious profession…
PyFi: Toward Pyramid-like Financial Image Understanding for VLMs via Adversarial Agents
Yuqun Zhang, Yuxuan Zhao, Sijia Chen
This paper proposes PyFi, a novel framework for pyramid-like financial image understanding that enables vision language models (VLMs) to reason through question chains in a progres…
CARES: Comprehensive Evaluation of Safety and Adversarial Robustness in Medical LLMs
Sijia Chen, Xiaomin Li, Mengxue Zhang +3
Large language models (LLMs) are increasingly deployed in medical contexts, raising critical concerns about safety, alignment, and susceptibility to adversarial manipulation. While…