55 citations · 55 across the 4 of their papers we have counts for
3 papers · 1 filter
VLAff: Vision-Language-Affordance Model for Unified Actionable Affordances
Jihoon Oh, Kento Kawaharazuka, Kei Okada
Learning manipulation skills from human videos is promising for scalable robot learning. However, the embodiment mismatch between humans and robots makes this challenging. One prom…
MEVION: Low-Cost Open-Source Data Collection System for Powerful and High-Speed Dual-Arm Manipulation
Kento Kawaharazuka, Yoshiki Obinata, Hirokazu Ishida +6
The global competition for developing robotic foundation models is intensifying. Among the data collection systems used for dual-arm robots, ALOHA is representative of being low-co…
Vision-Language-Action Models for Robotics: A Review Towards Real-World Applications
Kento Kawaharazuka, Jihoon Oh, Jun Yamada +2
Amid growing efforts to leverage advances in large language models (LLMs) and vision-language models (VLMs) for robotics, Vision-Language-Action (VLA) models have recently gained s…