1 paper · 1 filter
Jaymari Chua, Chen Wang, Liming Zhu +1
As reasoning agents become increasingly complex, aligning their underlying reasoning and decision-making processes with human conceptual models is a challenge for AI security and s…