11 citations · 13 across the 4 of their papers we have counts for
4 papers
Cross-modal Consistency Learning with Fine-grained Fusion Network for Multimodal Fake News Detection
Jun Li, Yi Bin, Jie Zou +2
Previous studies on multimodal fake news detection have observed the mismatch between text and images in the fake news and attempted to explore the consistency of multimodal news b…
Faster Video Moment Retrieval with Point-Level Supervision
Xun Jiang, Zailei Zhou, Xing Xu +3
Video Moment Retrieval (VMR) aims at retrieving the most relevant events from an untrimmed video with natural language queries. Existing VMR methods suffer from two defects: (1) ma…
Learning Semantic-Aware Knowledge Guidance for Low-Light Image Enhancement
Yuhui Wu, Chen Pan, Guoqing Wang +4
Low-light image enhancement (LLIE) investigates how to improve illumination and produce normal-light images. The majority of existing methods improve low-light images via a global…
ScanERU: Interactive 3D Visual Grounding based on Embodied Reference Understanding
Ziyang Lu, Yunqiang Pei, Guoqing Wang +3
Aiming to link natural language descriptions to specific regions in a 3D scene represented as 3D point clouds, 3D visual grounding is a very fundamental task for human-robot intera…