5 papers
Robot-Factored World Models via Robot Rendering
Byungjun Kim, Taeksoo Kim, Hyunsoo Cha +1
Action-conditioned video world models predict future observations from an initial observation and an action signal. In robotics, actions influence future observations through two d…
AutoDex: An Automated Real-World System for Dexterous Grasping Data Collection
Mingi Choi, Gunhee Kim, Jisoo Kim +4
Learning robust dexterous grasping requires real-world data that records the physical outcomes of grasp attempts. Such data is hard to obtain at scale: teleoperation yields valid p…
ZeroDex: Zero-Shot Long-Horizon Dexterous Manipulation via Multi-View 3D-Grounded VLM Reasoning
Jisoo Kim, Sangwon Baik, Taeksoo Kim +4
We present ZeroDex, a zero-shot framework for long-horizon dexterous manipulation that grounds language instructions into executable 3D task plans from calibrated multi-view RGB im…
Target-Aware Video Diffusion Models
Taeksoo Kim, Hanbyul Joo
We present a target-aware video diffusion model that generates videos from an input image, in which an actor interacts with a specified target while performing a desired action. Th…
Dexterous World Models
Byungjun Kim, Taeksoo Kim, Junyoung Lee +1
Recent progress in 3D reconstruction has made it easy to create realistic digital twins from everyday environments. However, current digital twins remain largely static and are lim…