1 paper
Terry Yue Zhuo, Yaqing Liao, Yuecheng Lei +5
We introduce ViLPAct, a novel vision-language benchmark for human activity planning. It is designed for a task where embodied AI agents can reason and forecast future actions of hu…