5 papers
See Selectively, Act Adaptively: Dual-Level Structural Decomposition for Bimanual Robot Manipulation
Yoon-Ji Choi, Young-Chae Son, Soo-Chul Lim
In bimanual robotic manipulation, task-relevant visual information varies with the task stage and context, while the interaction of the two arms shifts between independent and coor…
Pixel2Catch: Multi-Agent Sim-to-Real Transfer for Agile Manipulation with a Single RGB Camera
Seongyong Kim, Junhyeon Cho, Kang-Won Lee +1
To catch a thrown object, a robot must be able to perceive the object's motion and generate control actions in a timely manner. Rather than explicitly estimating the object's 3D po…
Selective Perception for Robot: Task-Aware Attention in Multimodal VLA
Young-Chae Son, Jung-Woo Lee, Yoon-Ji Choi +2
In robotics, Vision-Language-Action (VLA) models that integrate diverse multimodal signals from multi-view inputs have emerged as an effective approach. However, most prior work ad…
MemEIC: A Step Toward Continual and Compositional Knowledge Editing
Jin Seong, Jiyun Park, Wencke Liermann +5
The dynamic nature of information necessitates continuously updating large vision-language models (LVLMs). While recent knowledge editing techniques hint at promising directions, t…
DexTouch: Learning to Seek and Manipulate Objects with Tactile Dexterity
Kang-Won Lee, Yuzhe Qin, Xiaolong Wang +1
The sense of touch is an essential ability for skillfully performing a variety of tasks, providing the capacity to search and manipulate objects without relying on visual informati…