activity
20162020
most citedTAP: Text-Aware Pre-training for Text-VQA and Text-Caption

19 citations · 19 across the 2 of their papers we have counts for

collaborators

6 papers

cs.CV202019 cited

TAP: Text-Aware Pre-training for Text-VQA and Text-Caption

Zhengyuan Yang, Yijuan Lu, Jianfeng Wang +6

In this paper, we propose Text-Aware Pre-training (TAP) for Text-VQA and Text-Caption tasks. These two tasks aim at reading and understanding scene text in images for question answ…

cs.CV2018

RePr: Improved Training of Convolutional Filters

Aaditya Prakash, James Storer, Dinei Florencio +1

A well-trained Convolutional Neural Network can easily be pruned without significant loss of performance. This is because of unnecessary overlap in the features captured by the net…

cs.CV2018

A Fusion Framework for Camouflaged Moving Foreground Detection in the Wavelet Domain

Shuai Li, Dinei Florencio, Wanqing Li +2

Detecting camouflaged moving foreground objects has been known to be difficult due to the similarity between the foreground objects and the background. Conventional methods cannot…

cs.CL2018

Deep Learning Based Speech Beamforming

Kaizhi Qian, Yang Zhang, Shiyu Chang +3

Multi-channel speech enhancement with ad-hoc sensors has been a challenging task. Speech model guided beamforming algorithms are able to recover natural sounding speech, but the sp…

cs.CV2017

Foreground Detection in Camouflaged Scenes

Shuai Li, Dinei Florencio, Yaqin Zhao +2

Foreground detection has been widely studied for decades due to its importance in many practical applications. Most of the existing methods assume foreground and background show vi…

cs.SD2016

Speech Enhancement In Multiple-Noise Conditions using Deep Neural Networks

Anurag Kumar, Dinei Florencio

In this paper we consider the problem of speech enhancement in real-world like conditions where multiple noises can simultaneously corrupt speech. Most of the current literature on…