64 citations · 105 across the 13 of their papers we have counts for
14 papers
Audio-visual Representation Learning for Anomaly Events Detection in Crowds
Junyu Gao, Maoguo Gong, Xuelong Li
In recent years, anomaly events detection in crowd scenes attracts many researchers' attention, because of its importance to public safety. Existing methods usually exploit visual…
Self-adaptive Multi-task Particle Swarm Optimization
Xiaolong Zheng, Deyun Zhou, Na Li +3
Multi-task optimization (MTO) studies how to simultaneously solve multiple optimization problems for the purpose of obtaining better performance on each problem. Over the past few…
Hierarchical Multimodal Transformer to Summarize Videos
Bin Zhao, Maoguo Gong, Xuelong Li
Although video summarization has achieved tremendous success benefiting from Recurrent Neural Networks (RNN), RNN-based methods neglect the global dependencies and multi-hop relati…
Congested Crowd Instance Localization with Dilated Convolutional Swin Transformer
Junyu Gao, Maoguo Gong, Xuelong Li
Crowd localization is a new computer vision task, evolved from crowd counting. Different from the latter, it provides more precise location information for each instance, not just…
Towards Explainable Multi-Party Learning: A Contrastive Knowledge Sharing Framework
Yuan Gao, Jiawei Li, Maoguo Gong +2
Multi-party learning provides solutions for training joint models with decentralized data under legal and practical constraints. However, traditional multi-party learning approache…
AudioVisual Video Summarization
Bin Zhao, Maoguo Gong, Xuelong Li
Audio and vision are two main modalities in video data. Multimodal learning, especially for audiovisual learning, has drawn considerable attention recently, which can boost the per…