6 citations · 14 across the 7 of their papers we have counts for
7 papers · 1 filter
Improving Scene Graph Generation with Superpixel-Based Interaction Learning
Jingyi Wang, Can Zhang, Jinfa Huang +2
Recent advances in Scene Graph Generation (SGG) typically model the relationships among entities utilizing box-level features from pre-defined detectors. We argue that an overlooke…
Visual Information Extraction in the Wild: Practical Dataset and End-to-end Solution
Jianfeng Kuang, Wei Hua, Dingkang Liang +4
Visual information extraction (VIE), which aims to simultaneously perform OCR and information extraction in a unified framework, has drawn increasing attention due to its essential…
Scene Consistency Representation Learning for Video Scene Segmentation
Haoqian Wu, Keyu Chen, Yanan Luo +5
A long-term video, such as a movie or TV show, is composed of various scenes, each of which represents a series of shots sharing the same semantic story. Spotting the correct scene…
The Devil is in the Frequency: Geminated Gestalt Autoencoder for Self-Supervised Visual Pre-Training
Hao Liu, Xinghua Jiang, Xin Li +3
The self-supervised Masked Image Modeling (MIM) schema, following "mask-and-reconstruct" pipeline of recovering contents from masked image, has recently captured the increasing int…
Knowledge Mining with Scene Text for Fine-Grained Recognition
Hao Wang, Junchao Liao, Tianheng Cheng +5
Recently, the semantics of scene text has been proven to be essential in fine-grained image classification. However, the existing methods mainly exploit the literal meaning of scen…
PuzzleNet: Scene Text Detection by Segment Context Graph Learning
Hao Liu, Antai Guo, Deqiang Jiang +2
Recently, a series of decomposition-based scene text detection methods has achieved impressive progress by decomposing challenging text regions into pieces and linking them in a bo…