11 papers
Back to the Familiar Future: Failure Recovery for VLA Policies via Pre-Imagined Milestone Selection
Suyeon Shin, Juwon Kim, Hyeonbin Park +4
Vision-language-action (VLA) policies can deviate from nominal trajectories during manipulation, even when tasks remain physically feasible. Recovering from these deviations is cha…
DBMovi-GS: Dynamic View Synthesis from Blurry Monocular Video via Sparse-Controlled Gaussian Splatting
Yeon-Ji Song, Jaein Kim, Byung-Ju Kim +1
Novel view synthesis is a task of generating scenes from unseen perspectives; however, synthesizing dynamic scenes from blurry monocular videos remains an unresolved challenge that…
CLIP-RT: Learning Language-Conditioned Robotic Policies from Natural Language Supervision
Gi-Cheon Kang, Junghyun Kim, Kyuhwan Shim +2
Teaching robots desired skills in real-world environments remains challenging, especially for non-experts. A key bottleneck is that collecting robotic data often requires expertise…
Zero-Shot Vision-and-Language Navigation with Collision Mitigation in Continuous Environment
Seongjun Jeong, Gi-Cheon Kang, Joochan Kim +1
We propose the zero-shot Vision-and-Language Navigation with Collision Mitigation (VLN-CM), which takes these considerations. VLN-CM is composed of four modules and predicts the di…
HAPFI: History-Aware Planning based on Fused Information
Sujin Jeon, Suyeon Shin, Byoung-Tak Zhang
Embodied Instruction Following (EIF) is a task of planning a long sequence of sub-goals given high-level natural language instructions, such as "Rinse a slice of lettuce and place…
Surface-Based Visibility-Guided Uncertainty for Continuous Active 3D Neural Reconstruction
Hyunseo Kim, Hyeonseo Yang, Taekyung Kim +4
View selection is critical in active 3D neural reconstruction as it impacts the contents of training set and resulting final output quality. Recent view selection strategies emphas…