2 citations · 4 across the 3 of their papers we have counts for
3 papers
cs.CV2024★ 1 cited
ACTRESS: Active Retraining for Semi-supervised Visual Grounding
Weitai Kang, Mengxue Qu, Yunchao Wei +1
Semi-Supervised Visual Grounding (SSVG) is a new challenge for its sparse labeled data with the need for multimodel understanding. A previous study, RefTeacher, makes the first att…
cs.CV2023★ 2 cited
RIO: A Benchmark for Reasoning Intention-Oriented Objects in Open Environments
Mengxue Qu, Yu Wu, Wu Liu +4
Intention-oriented object detection aims to detect desired objects based on specific intentions or requirements. For instance, when we desire to "lie down and rest", we instinctive…
cs.CV2022★ 1 cited
SiRi: A Simple Selective Retraining Mechanism for Transformer-based Visual Grounding
Mengxue Qu, Yu Wu, Wu Liu +5
In this paper, we investigate how to achieve better visual grounding with modern vision-language transformers, and propose a simple yet powerful Selective Retraining (SiRi) mechani…