237 citations · 1.1k across the 44 of their papers we have counts for
65 papers
Relation Regularized Scene Graph Generation
Yuyu Guo, Lianli Gao, Jingkuan Song +4
Scene graph generation (SGG) is built on top of detected objects to predict object pairwise visual relations for describing the image content abstraction. Existing works have revea…
Audio-visual Representation Learning for Anomaly Events Detection in Crowds
Junyu Gao, Maoguo Gong, Xuelong Li
In recent years, anomaly events detection in crowd scenes attracts many researchers' attention, because of its importance to public safety. Existing methods usually exploit visual…
Unsupervised Domain Adaptive Learning via Synthetic Data for Person Re-identification
Qi Wang, Sikai Bai, Junyu Gao +2
Person re-identification (re-ID) has gained more and more attention due to its widespread applications in intelligent video surveillance. Unfortunately, the mainstream deep learnin…
LDC-Net: A Unified Framework for Localization, Detection and Counting in Dense Crowds
Qi wang, Tao Han, Junyu Gao +2
The rapid development in visual crowd analysis shows a trend to count people by positioning or even detecting, rather than simply summing a density map. It also enlightens us back…
Hierarchical Multimodal Transformer to Summarize Videos
Bin Zhao, Maoguo Gong, Xuelong Li
Although video summarization has achieved tremendous success benefiting from Recurrent Neural Networks (RNN), RNN-based methods neglect the global dependencies and multi-hop relati…
Congested Crowd Instance Localization with Dilated Convolutional Swin Transformer
Junyu Gao, Maoguo Gong, Xuelong Li
Crowd localization is a new computer vision task, evolved from crowd counting. Different from the latter, it provides more precise location information for each instance, not just…