1 paper
Jaehwan Jeong, Evelyn Zhu, Jinying Lin +5
Vision-Language-Action (VLA) models have demonstrated strong potential for predicting semantic actions in navigation tasks, demonstrating the ability to reason over complex linguis…