From the 1 of 15 linked papers with an AI index.
1 citations · 1 across the 7 of their papers we have counts for
3 papers · 1 filter
S2A2: Audio-Visual Imitation Learning for Manipulation Tasks Using Acoustic Spatial Information
Kaneyoshi Hiratsuka, Benjamin Yen, Ryosuke Kojima
The paper presents acoustic-aware manipulation tasks where robots use sound cues to locate and identify objects, and introduces the S2A2 multimodal imitation learning framework tha…
From Sign Language Generation to Humanoid Execution: Vision-Language Guided Retargeting with Collision Mitigation
Nabeela Khan, Bowen Wu, Runwu Shi +5
Recent sign language generation (SLG) systems increasingly output dense 3D body representations, which better preserve full-body kinematics and geometry for downstream embodiment o…
Single-Microphone-Based Sound Source Localization for Mobile Robots in Reverberant Environments
Jiang Wang, Runwu Shi, Benjamin Yen +2
Accurately estimating sound source positions is crucial for robot audition. However, existing sound source localization methods typically rely on a microphone array with at least t…