activity
20152021
most citedThe Cross-Depiction Problem: Computer Vision Algorithms for Recognising Objects in Artwork and in Photographs

31 citations · 136 across the 10 of their papers we have counts for

collaborators
Showing cs.CVShow all

12 papers · 1 filter

cs.CV2021

Towards Accurate Text-based Image Captioning with Content Diversity Exploration

Guanghui Xu, Shuaicheng Niu, Mingkui Tan +3

Text-based image captioning (TextCap) which aims to read and reason images with texts is crucial for a machine to understand a detailed and complex scene environment, considering t…

cs.CV202117 cited

Non-Salient Region Object Mining for Weakly Supervised Semantic Segmentation

Yazhou Yao, Tao Chen, Guosen Xie +5

Semantic segmentation aims to classify every pixel of an input image. Considering the difficulty of acquiring dense labels, researchers have recently been resorting to weak labels…

cs.CV20216 cited

Jo-SRC: A Contrastive Approach for Combating Noisy Labels

Yazhou Yao, Zeren Sun, Chuanyi Zhang +4

Due to the memorization effect in Deep Neural Networks (DNNs), training with noisy labels usually results in inferior model performance. Existing state-of-the-art methods primarily…

cs.CV2019

Show, Price and Negotiate: A Negotiator with Online Value Look-Ahead

Amin Parvaneh, Ehsan Abbasnejad, Qi Wu +2

Negotiation, as an essential and complicated aspect of online shopping, is still challenging for an intelligent agent. To that end, we propose the Price Negotiator, a modular deep…

cs.CV20191 cited

You Only Look & Listen Once: Towards Fast and Accurate Visual Grounding

Chaorui Deng, Qi Wu, Guanghui Xu +4

Visual Grounding (VG) aims to locate the most relevant region in an image, based on a flexible natural language query but not a pre-defined label, thus it can be a more useful tech…

cs.CV201717 cited

Asking the Difficult Questions: Goal-Oriented Visual Question Generation via Intermediate Rewards

Junjie Zhang, Qi Wu, Chunhua Shen +3

Despite significant progress in a variety of vision-and-language problems, developing a method capable of asking intelligent, goal-oriented questions about images is proven to be a…