3 papers
cs.AI2026
GDPevo: Evaluating Agent Self-Evolution on Real Business Tasks
Leijun Zhou, Zhihao Liu, Xiang Qu +9
Agent self-evolution updates an agent's persistent state from prior experience and reuses it to solve related tasks more effectively. Evaluating self-evolution is difficult: existi…
cs.LG2026
Discovering Decoupled Functional Modules in Large Language Models
Yanke Yu, Jin Li, Ying Sun +3
Understanding the internal functional organization of Large Language Models (LLMs) is crucial for improving their trustworthiness and performance. However, how LLMs organize differ…
cs.IR2025
Improving Recommendation Fairness without Sensitive Attributes Using Multi-Persona LLMs
Haoran Xin, Ying Sun, Chao Wang +3
Despite the success of recommender systems in alleviating information overload, fairness issues have raised concerns in recent years, potentially leading to unequal treatment for c…