13 citations · 23 across the 10 of their papers we have counts for
Showing cs.AIShow all
2 papers · 1 filter
cs.AI2026
Preference Data Selection for Mitigating the Alignment Tax in Large Language Models
Minsu Kim, Jianxun Lian, Xing Xie +1
Aligning large language models to human preferences is crucial for real-world deployment but frequently incurs an alignment tax, leading to the catastrophic forgetting of pre-train…
cs.AI2026
To Think or Not To Think, That is The Question for Large Reasoning Models in Theory of Mind Tasks
Nanxu Gong, Haotian Li, Sixun Dong +3
Theory of Mind (ToM) assesses whether models can infer hidden mental states such as beliefs, desires, and intentions, which is essential for natural social interaction. Although re…