1 paper · 1 filter
Yuhan Wang, Zhengxi Lu, Yuchen Yan +6
Research planning is the decisive capability of AI scientists. Yet a research plan admits no verifiable answer, so reinforcement learning lacks the environment it requires: tasks p…