5 citations · 6 across the 10 of their papers we have counts for
5 papers · 1 filter
AI Evaluation Should Require Standardized Item-Level Data Releases
Han Jiang, Susu Zhang, Dongyao Zhu +6
This position paper argues that standardized item-level benchmark data should become the default infrastructure for AI evaluation. Current evaluations suffer from underspecified it…
On the Dynamics of Multi-Agent LLM Communities Driven by Value Diversity
Muhua Huang, Qinlin Zhao, Xiaoyuan Yi +1
As Large Language Models (LLM) based multi-agent systems become increasingly prevalent, the collective behaviors, e.g., collective intelligence, of such artificial communities have…
The Morality of Probability: How Implicit Moral Biases in LLMs May Shape the Future of Human-AI Symbiosis
Eoin O'Doherty, Nicole Weinrauch, Andrew Talone +4
Artificial intelligence (AI) is advancing at a pace that raises urgent questions about how to align machine decision-making with human moral values. This working paper investigates…
The Incomplete Bridge: How AI Research (Mis)Engages with Psychology
Han Jiang, Pengda Wang, Xiaoyuan Yi +2
Social sciences have accumulated a rich body of theories and methodologies for investigating the human mind and behaviors, while offering valuable insights into the design and unde…
Image Inspired Poetry Generation in XiaoIce
Wen-Feng Cheng, Chao-Chung Wu, Ruihua Song +3
Vision is a common source of inspiration for poetry. The objects and the sentimental imprints that one perceives from an image may lead to various feelings depending on the reader.…