From the 1 of 10 linked papers with an AI index.
10 papers
Group Alignment-Induced Sycophancy: A Two-Sided Evaluation of Steerable Pluralistic Alignment
Haokai Zhao, Yunze Xiao, Weihao Xuan +3
Group alignment adapts a language model to a demographic group to produce responses that reflect the group's opinions, values, and preferences. Sycophancy, a well-documented by-pro…
ExpressionCueLens: A Cross-Cultural Analysis of Human-AI Companion Conversations on Social Media
Lynnette Hui Xian Ng, Yunze Xiao, Lionel Z. Wang +2
The paper presents ExpressionCueLens, a framework for categorizing anthropomorphic expressions in human‑AI companion conversations, and uses it to compare how Reddit and XiaoHongSh…
MixSD: Mixed Contextual Self-Distillation for Knowledge Injection
Jiarui Liu, Lechen Zhang, Yongjin Yang +5
Supervised fine-tuning (SFT) is widely used to inject new knowledge into language models, but it often degrades pretrained capabilities such as reasoning and general-domain perform…
Knowledge Index of Noah's Ark
Sheng Jin, Minghao Liu, Yunze Xiao +24
Knowledge benchmarks for LLMs face three issues: scaling-driven designs that do not operationalize disciplinary representativeness; flat-payment annotation that permits lazy consen…
Every Act Has Its Price: Compressed Moral Composition in Frontier LLMs
Weijia Zhang, Ruiqi Chen, Yunze Xiao +1
Existing LLM moral benchmarks usually ask which isolated moral act, value, or foundation a model prefers. This is useful but incomplete. Realistic judgments often require a model t…
The Chameleon's Limit: Investigating Persona Collapse and Homogenization in Large Language Models
Yunze Xiao, Vivienne J. Zhang, Chenghao Yang +3
Applications based on large language models (LLMs), such as multi-agent simulations, require population diversity among agents. We identify a pervasive failure mode we term \emph{P…