From the 1 of 3 linked papers with an AI index.
1 paper · 1 filter
Xiaoyi Bao, Yuanzhen Xie, Yunzhi Tan +5
The paper presents SkillMentor, a reinforcement‑learning trained mentor that enables large language model agents to learn how to identify and diagnose their own blind‑spot failures…