35 citations · 37 across the 2 of their papers we have counts for
3 papers
cs.CV2021
Chop Chop BERT: Visual Question Answering by Chopping VisualBERT's Heads
Chenyu Gao, Qi Zhu, Peng Wang +1
Vision-and-Language (VL) pre-training has shown great potential on many related downstream tasks, such as Visual Question Answering (VQA), one of the most popular problems in the V…
cs.CV2020★ 2 cited
Simple is not Easy: A Simple Strong Baseline for TextVQA and TextCaps
Qi Zhu, Chenyu Gao, Peng Wang +1
Texts appearing in daily scenes that can be recognized by OCR (Optical Character Recognition) tools contain significant information, such as street name, product brand and prices.…
cs.CV2019★ 35 cited
C^3 Framework: An Open-source PyTorch Code for Crowd Counting
Junyu Gao, Wei Lin, Bin Zhao +3
This technical report attempts to provide efficient and solid kits addressed on the field of crowd counting, which is denoted as Crowd Counting Code Framework (CF). The contrib…