1 paper
Jooyoung Kim, Wonje Choi, Younguk Song +1
Recent advances in Vision-Language Models (VLMs) have enabled video-instructed robotic programming, allowing agents to interpret video demonstrations and generate executable contro…