2 papers
cs.LG2025
Cross-modal Causal Relation Alignment for Video Question Grounding
Weixing Chen, Yang Liu, Binglin Chen +3
Video question grounding (VideoQG) requires models to answer the questions and simultaneously infer the relevant video segments to support the answers. However, existing VideoQG me…
cs.CV2024
ODMixer: Fine-grained Spatial-temporal MLP for Metro Origin-Destination Prediction
Yang Liu, Binglin Chen, Yongsen Zheng +3
Metro Origin-Destination (OD) prediction is a crucial yet challenging spatial-temporal prediction task in urban computing, which aims to accurately forecast cross-station ridership…