3 papers
cs.MA2026
When Collaboration Becomes a Trigger: Collective Evidence-Threshold Backdoors in Multi-Agent Systems
Jia-Hao Xiao, Lei Feng, Min-Ling Zhang
LLM-based multi-agent systems (MAS) extend LLM capabilities through iterative communication and shared contexts. However, this collaboration introduces a vulnerability: backdoor be…
cs.CL2025
Rethinking Toxicity Evaluation in Large Language Models: A Multi-Label Perspective
Zhiqiang Kou, Junyang Chen, Xin-Qiang Cai +8
Large language models (LLMs) have achieved impressive results across a range of natural language processing tasks, but their potential to generate harmful content has raised seriou…
cs.CV2025
What Makes "Good" Distractors for Object Hallucination Evaluation in Large Vision-Language Models?
Ming-Kun Xie, Jia-Hao Xiao, Gang Niu +4
Large Vision-Language Models (LVLMs), empowered by the success of Large Language Models (LLMs), have achieved impressive performance across domains. Despite the great advances in L…