Showing cs.AIShow all
3 papers · 1 filter
cs.AI2026
PEAR: Permutation-Equivariant Adaptive Routing Multi-Agent Debate
Yang Feng, Ziwei Xu, Xia Hu +1
Multi-agent debate improves the reliability of large language models (LLMs) through iterative peer critiques. However, fixed topologies often introduce persistent positional biases…
cs.AI2025
Bullying the Machine: How Personas Increase LLM Vulnerability
Ziwei Xu, Udit Sanghi, Mohan Kankanhalli
Large Language Models (LLMs) are increasingly deployed in interactions where they are prompted to adopt personas. This paper investigates whether such persona conditioning affects…
cs.AI2025
Strong Preferences Affect the Robustness of Preference Models and Value Alignment
Ziwei Xu, Mohan Kankanhalli
Value alignment, which aims to ensure that large language models (LLMs) and other AI agents behave in accordance with human values, is critical for ensuring safety and trustworthin…