119 citations · 266 across the 9 of their papers we have counts for
11 papers
Renmin University of China at TRECVID 2022: Improving Video Search by Feature Fusion and Negation Understanding
Xirong Li, Aozhu Chen, Ziyue Wang +4
We summarize our TRECVID 2022 Ad-hoc Video Search (AVS) experiments. Our solution is built with two new techniques, namely Lightweight Attentional Feature Fusion (LAFF) for combini…
Lesion Localization in OCT by Semi-Supervised Object Detection
Yue Wu, Yang Zhou, Jianchun Zhao +4
Over 300 million people worldwide are affected by various retinal diseases. By noninvasive Optical Coherence Tomography (OCT) scans, a number of abnormal structural changes in the…
DRAG: Dynamic Region-Aware GCN for Privacy-Leaking Image Detection
Guang Yang, Juan Cao, Qiang Sheng +3
The daily practice of sharing images on social media raises a severe issue about privacy leakage. To address the issue, privacy-leaking image detection is studied recently, with th…
Deepfake Network Architecture Attribution
Tianyun Yang, Ziyao Huang, Juan Cao +2
With the rapid progress of generation technology, it has become necessary to attribute the origin of fake images. Existing works on fake image attribution perform multi-class class…
Reading-strategy Inspired Visual Representation Learning for Text-to-Video Retrieval
Jianfeng Dong, Yabing Wang, Xianke Chen +4
This paper aims for the task of text-to-video retrieval, where given a query in the form of a natural-language sentence, it is asked to retrieve videos which are semantically relev…
Multi-Modal Multi-Instance Learning for Retinal Disease Recognition
Xirong Li, Yang Zhou, Jie Wang +5
This paper attacks an emerging challenge of multi-modal retinal disease recognition. Given a multi-modal case consisting of a color fundus photo (CFP) and an array of OCT B-scan im…