2.2k citations
- Peking UniversityCN157 papers
- Institute of High Energy PhysicsCN150 papers
- Tsinghua UniversityCN148 papers
- University of Science and Technology of ChinaCN143 papers
- Nanjing UniversityCN139 papers
- Shandong UniversityCN132 papers
- Sichuan UniversityCN125 papers
- Zhejiang UniversityCN123 papers
- Sun Yat-sen UniversityCN122 papers
- China Center of Advanced Science and TechnologyCN120 papers
- Nanjing Normal UniversityCN120 papers
- Wuhan UniversityCN120 papers
29 papers · 1 filter
Video Understanding: From Geometry and Semantics to Unified Models
Zhaochong An, Zirui Li, Mingqiao Ye +9
Video understanding aims to enable models to perceive, reason about, and interact with the dynamic visual world. In contrast to image understanding, video understanding inherently…
Evaluating SAM2 for Video Semantic Segmentation
Syed Hesham Syed Ariff, Yun Liu, Guolei Sun +4
The Segmentation Anything Model 2 (SAM2) has proven to be a powerful foundation model for promptable visual object segmentation in both images and videos, capable of storing object…
Bridging Visual Affective Gap: Borrowing Textual Knowledge by Learning from Noisy Image-Text Pairs
Daiqing Wu, Dongbao Yang, Yu Zhou +1
Visual emotion recognition (VER) is a longstanding field that has garnered increasing attention with the advancement of deep neural networks. Although recent studies have achieved…
GAPNet: A Lightweight Framework for Image and Video Salient Object Detection via Granularity-Aware Paradigm
Yu-Huan Wu, Wei Liu, Zi-Xuan Zhu +3
Recent salient object detection (SOD) models predominantly rely on heavyweight backbones, incurring substantial computational cost and hindering their practical application in vari…
A Reverse Causal Framework to Mitigate Spurious Correlations for Debiasing Scene Graph Generation
Shuzhou Sun, Li Liu, Tianpeng Liu +4
Existing two-stage Scene Graph Generation (SGG) frameworks typically incorporate a detector to extract relationship features and a classifier to categorize these relationships; the…
DEYOLO: Dual-Feature-Enhancement YOLO for Cross-Modality Object Detection
Yishuo Chen, Boran Wang, Xinyu Guo +4
Object detection in poor-illumination environments is a challenging task as objects are usually not clearly visible in RGB images. As infrared images provide additional clear edge…