3 papers
cs.CV2025
IMoRe: Implicit Program-Guided Reasoning for Human Motion Q&A
Chen Li, Chinthani Sugandhika, Yeo Keat Ee +5
Existing human motion Q\&A methods rely on explicit program execution, where the requirement for manually defined functional modules may limit the scalability and adaptability. To…
cs.CV2025
Neuro Symbolic Knowledge Reasoning for Procedural Video Question Answering
Basura Fernando, Thanh-Son Nguyen, Hong Yang +3
In this work we present Knowledge Module Learning (KML) to understand and reason over procedural tasks that requires models to learn structured and compositional procedural knowled…
cs.CV2024
Training-Free Action Recognition and Goal Inference with Dynamic Frame Selection
Ee Yeo Keat, Zhang Hao, Alexander Matyasko +1
We introduce VidTFS, a Training-free, open-vocabulary video goal and action inference framework that combines the frozen vision foundational model (VFM) and large language model (L…