5 papers
MulCPred: Learning Multi-modal Concepts for Explainable Pedestrian Action Prediction
Yan Feng, Alexander Carballo, Keisuke Fujii +3
Pedestrian action prediction is of great significance for many applications such as autonomous driving. However, state-of-the-art methods lack explainability to make trustworthy pr…
DRUformer: Enhancing the driving scene Important object detection with driving relationship self-understanding
Yingjie Niu, Ming Ding, Keisuke Fujii +3
Traffic accidents frequently lead to fatal injuries, contributing to over 50 million deaths until 2023. To mitigate driving hazards and ensure personal safety, it is crucial to ass…
Automatic Edge Error Judgment in Figure Skating Using 3D Pose Estimation from a Monocular Camera and IMUs
Ryota Tanaka, Tomohiro Suzuki, Kazuya Takeda +1
Automatic evaluating systems are fundamental issues in sports technologies. In many sports, such as figure skating, automated evaluating methods based on pose estimation have been…
Runner re-identification from single-view running video in the open-world setting
Tomohiro Suzuki, Kazushi Tsutsui, Kazuya Takeda +1
In many sports, player re-identification is crucial for automatic video processing and analysis. However, most of the current studies on player re-identification in multi- or singl…
Compositional Semantics for Open Vocabulary Spatio-semantic Representations
Robin Karlsson, Francisco Lepe-Salazar, Kazuya Takeda
Vision-language models (VLMs) transform environment percepts into vision-language semantics interpretable by LLMs. However, completing complex tasks often requires reasoning about…