1 paper
Pei Liu, Qingtian Ning, Xinyan Lu +6
The pursuit of autonomous agents capable of temporally coherent planning is hindered by a fundamental flaw in current vision-language models (VLMs): they lack cognitive inertia. Op…