11 citations · 13 across the 3 of their papers we have counts for
3 papers
cs.CV2023
AutoShot: A Short Video Dataset and State-of-the-Art Shot Boundary Detection
Wentao Zhu, Yufang Huang, Xiufeng Xie +5
The short-form videos have explosive popularity and have dominated the new social media trends. Prevailing short-video platforms,~\textit{e.g.}, Kuaishou (Kwai), TikTok, Instagram…
cs.RO2023★ 2 cited
LP-SLAM: Language-Perceptive RGB-D SLAM system based on Large Language Model
Weiyi Zhang, Yushi Guo, Liting Niu +6
Simultaneous localization and mapping (SLAM) is a critical technology that enables autonomous robots to be aware of their surrounding environment. With the development of deep lear…
cs.CV2021★ 11 cited
A Bilingual, OpenWorld Video Text Dataset and End-to-end Video Text Spotter with Transformer
Weijia Wu, Yuanqiang Cai, Debing Zhang +5
Most existing video text spotting benchmarks focus on evaluating a single language and scenario with limited data. In this work, we introduce a large-scale, Bilingual, Open World V…