5 papers
Touch2Insert: Zero-Shot Peg Insertion by Touching Intersections of Peg and Hole
Masaru Yajima, Yuma Shin, Rei Kawakami +2
Reliable insertion of industrial connectors remains a central challenge in robotics, requiring sub-millimeter precision under uncertainty and often without full visual access. Visi…
Teach Me Sign: Stepwise Prompting LLM for Sign Language Production
Zhaoyi An, Rei Kawakami
Large language models, with their strong reasoning ability and rich knowledge, have brought revolution to many tasks of AI, but their impact on sign language generation remains lim…
A Unified Transformer-Based Framework with Pretraining For Whole Body Grasping Motion Generation
Edward Effendy, Kuan-Wei Tseng, Rei Kawakami
Accepted in the ICIP 2025 We present a novel transformer-based framework for whole-body grasping that addresses both pose generation and motion infilling, enabling realistic and st…
Anomaly Object Segmentation with Vision-Language Models for Steel Scrap Recycling
Daichi Tanaka, Takumi Karasawa, Shu Takenouchi +1
Recycling steel scrap can reduce carbon dioxide (CO2) emissions from the steel industry. However, a significant challenge in steel scrap recycling is the inclusion of impurities ot…
Zero-Shot Peg Insertion: Identifying Mating Holes and Estimating SE(2) Poses with Vision-Language Models
Masaru Yajima, Kei Ota, Asako Kanezaki +1
Achieving zero-shot peg insertion, where inserting an arbitrary peg into an unseen hole without task-specific training, remains a fundamental challenge in robotics. This task deman…