Showing cs.AIShow all
2 papers · 1 filter
cs.AI2026
Illusion of Alignment: Detecting Hidden Disagreement in Collaborative Dialogue
Kaiming Liu, Fuwen Luo, Ziyue Wang +6
Collaborative dialogue can end with apparent agreement while participants still differ on goals, assumptions, or execution plans, creating an \textbf{illusion of alignment (IoA)}.…
cs.AI2024
Towards Unified Alignment Between Agents, Humans, and Environment
Zonghan Yang, An Liu, Zijun Liu +11
The rapid progress of foundation models has led to the prosperity of autonomous agents, which leverage the universal capabilities of foundation models to conduct reasoning, decisio…