44 citations · 77 across the 10 of their papers we have counts for
7 papers · 1 filter
SkeletonMAE: Graph-based Masked Autoencoder for Skeleton Sequence Pre-training
Hong Yan, Yang Liu, Yushen Wei +3
Skeleton sequence representation learning has shown great advantages for action recognition due to its promising ability to model human joints and topology. However, the current me…
Spatio-Temporal Domain Awareness for Multi-Agent Collaborative Perception
Kun Yang, Dingkang Yang, Jingyu Zhang +6
Multi-agent collaborative perception as a potential application for vehicle-to-everything communication could significantly improve the perception performance of autonomous vehicle…
CausalVLR: A Toolbox and Benchmark for Visual-Linguistic Causal Reasoning
Yang Liu, Weixing Chen, Guanbin Li +1
We present CausalVLR (Causal Visual-Linguistic Reasoning), an open-source toolbox containing a rich set of state-of-the-art causal relation discovery and causal inference methods f…
Enhance Multi-Scale Spatial-Temporal Coherence for Configurable Video Anomaly Detection
Kai Cheng, Xinzhe Li, Lijuan Che
The development of unsupervised Video Anomaly Detection (VAD) relies on technologies in the field of signal processing. Since the anomaly is quite ambiguous and unbounded, differen…
VCD: Visual Causality Discovery for Cross-Modal Question Reasoning
Yang Liu, Ying Tan, Jingzhou Luo +1
Existing visual question reasoning methods usually fail to explicitly discover the inherent causal mechanism and ignore jointly modeling cross-modal event temporality and causality…
Exploring Interactions and Regulations in Collaborative Learning: An Interdisciplinary Multimodal Dataset
Yante Li, Yang Liu, KhÁnh Nguyen +4
Collaborative learning is an educational approach that enhances learning through shared goals and working together. Interaction and regulation are two essential factors related to…