1 citations · 1 across the 3 of their papers we have counts for
3 papers
cs.CV2024
ClickTrack: Towards Real-time Interactive Single Object Tracking
Kuiran Wang, Xuehui Yu, Wenwen Yu +5
Single object tracking(SOT) relies on precise object bounding box initialization. In this paper, we reconsidered the deficiencies in the current approaches to initializing single o…
cs.CV2024★ 1 cited
OmniParser: A Unified Framework for Text Spotting, Key Information Extraction and Table Recognition
Jianqiang Wan, Sibo Song, Wenwen Yu +6
Recently, visually-situated text parsing (VsTP) has experienced notable advancements, driven by the increasing demand for automated document understanding and the emergence of Gene…
cs.CV2023
Turning a CLIP Model into a Scene Text Spotter
Wenwen Yu, Yuliang Liu, Xingkui Zhu +3
We exploit the potential of the large-scale Contrastive Language-Image Pretraining (CLIP) model to enhance scene text detection and spotting tasks, transforming it into a robust ba…