8 citations · 18 across the 17 of their papers we have counts for
Showing cs.CVShow all
2 papers · 1 filter
cs.CV2024★ 1 cited
Open-Vocabulary Action Localization with Iterative Visual Prompting
Naoki Wake, Atsushi Kanehira, Kazuhiro Sasabuchi +2
Video action localization aims to find the timings of specific actions from a long video. Although existing learning-based approaches have been successful, they require annotating…
cs.CV2020★ 2 cited
Understanding Action Sequences based on Video Captioning for Learning-from-Observation
Iori Yanokura, Naoki Wake, Kazuhiro Sasabuchi +2
Learning actions from human demonstration video is promising for intelligent robotic systems. Extracting the exact section and re-observing the extracted video section in detail is…