3 papers
cs.CV2026
Anomalous Frame Detection by Grouping Frame Similarities between Two Videos Computed by Vision-Language Model to Extract Expert Workers' Unique Actions
Ryo Sakai, Yongpeng Cao, Nobutaka Kimura
Maintenance of critical infrastructures, such as railways and power plants, is essential for operational safety and reliability. However, the declining number of skilled maintenanc…
cs.CV2026
High-Speed Vision Improves Zero-Shot Semantic Understanding of Human Actions
Yongpeng Cao, Yuji Yamakawa
Understanding human actions from visual observations is essential for human--robot interaction, particularly when semantic interpretation of unfamiliar or hard-to-annotate actions…
cs.RO2026
SASI: Leveraging Sub-Action Semantics for Robust Early Action Recognition in Human-Robot Interaction
Yongpeng Cao, Masahiro Hirano, Hyuno Kim +1
Understanding human actions is critical for advancing behavior analysis in human-robot interaction. Particularly in tasks that demand quick and proactive feedback, robots must reco…