2 papers
cs.CV2025
Bidirectional Action Sequence Learning for Long-term Action Anticipation with Large Language Models
Yuji Sato, Yasunori Ishii, Takayoshi Yamashita
Video-based long-term action anticipation is crucial for early risk detection in areas such as automated driving and robotics. Conventional approaches extract features from past ac…
cs.CV2024
VDMA: Video Question Answering with Dynamically Generated Multi-Agents
Noriyuki Kugo, Tatsuya Ishibashi, Kosuke Ono +1
This technical report provides a detailed description of our approach to the EgoSchema Challenge 2024. The EgoSchema Challenge aims to identify the most appropriate responses to qu…