2 citations · 5 across the 11 of their papers we have counts for
6 papers · 1 filter
SeqCo-DETR: Sequence Consistency Training for Self-Supervised Object Detection with Transformers
Guoqiang Jin, Fan Yang, Mingshan Sun +7
Self-supervised pre-training and transformer-based networks have significantly improved the performance of object detection. However, most of the current self-supervised object det…
Generation-Guided Multi-Level Unified Network for Video Grounding
Xing Cheng, Xiangyu Wu, Dong Shen +2
Video grounding aims to locate the timestamps best matching the query description within an untrimmed video. Prevalent methods can be divided into moment-level and clip-level frame…
Spatial Attention and Syntax Rule Enhanced Tree Decoder for Offine Handwritten Mathematical Expression Recognition
Zihao Lin, Jinrong Li, Fan Yang +4
Offline Handwritten Mathematical Expression Recognition (HMER) has been dramatically advanced recently by employing tree decoders as part of the encoder-decoder method. Despite the…
Surrogate-assisted Multi-objective Neural Architecture Search for Real-time Semantic Segmentation
Zhichao Lu, Ran Cheng, Shihua Huang +3
The architectural advancements in deep neural networks have led to remarkable leap-forwards across a broad array of computer vision tasks. Instead of relying on human expertise, ne…
ReMix: A General and Efficient Framework for Multiple Instance Learning based Whole Slide Image Classification
Jiawei Yang, Hanbo Chen, Yu Zhao +4
Whole slide image (WSI) classification often relies on deep weakly supervised multiple instance learning (MIL) methods to handle gigapixel resolution images and slide-level labels.…
MSDU-net: A Multi-Scale Dilated U-net for Blur Detection
Fan Yang, Xiao Xiao
Blur detection is the separation of blurred and clear regions of an image, which is an important and challenging task in computer vision. In this work, we regard blur detection as…