31 citations · 136 across the 10 of their papers we have counts for
12 papers · 1 filter
Towards Accurate Text-based Image Captioning with Content Diversity Exploration
Guanghui Xu, Shuaicheng Niu, Mingkui Tan +3
Text-based image captioning (TextCap) which aims to read and reason images with texts is crucial for a machine to understand a detailed and complex scene environment, considering t…
Non-Salient Region Object Mining for Weakly Supervised Semantic Segmentation
Yazhou Yao, Tao Chen, Guosen Xie +5
Semantic segmentation aims to classify every pixel of an input image. Considering the difficulty of acquiring dense labels, researchers have recently been resorting to weak labels…
Jo-SRC: A Contrastive Approach for Combating Noisy Labels
Yazhou Yao, Zeren Sun, Chuanyi Zhang +4
Due to the memorization effect in Deep Neural Networks (DNNs), training with noisy labels usually results in inferior model performance. Existing state-of-the-art methods primarily…
Show, Price and Negotiate: A Negotiator with Online Value Look-Ahead
Amin Parvaneh, Ehsan Abbasnejad, Qi Wu +2
Negotiation, as an essential and complicated aspect of online shopping, is still challenging for an intelligent agent. To that end, we propose the Price Negotiator, a modular deep…
You Only Look & Listen Once: Towards Fast and Accurate Visual Grounding
Chaorui Deng, Qi Wu, Guanghui Xu +4
Visual Grounding (VG) aims to locate the most relevant region in an image, based on a flexible natural language query but not a pre-defined label, thus it can be a more useful tech…
Asking the Difficult Questions: Goal-Oriented Visual Question Generation via Intermediate Rewards
Junjie Zhang, Qi Wu, Chunhua Shen +3
Despite significant progress in a variety of vision-and-language problems, developing a method capable of asking intelligent, goal-oriented questions about images is proven to be a…