11 citations · 17 across the 5 of their papers we have counts for
7 papers · 1 filter
Mask-free OVIS: Open-Vocabulary Instance Segmentation without Manual Mask Annotations
Vibashan VS, Ning Yu, Chen Xing +5
Existing instance segmentation models learn task-specific information using manual mask annotations from base (training) categories. These mask annotations require tremendous human…
Robustness Evaluation of Transformer-based Form Field Extractors via Form Attacks
Le Xue, Mingfei Gao, Zeyuan Chen +2
We propose a novel framework to evaluate the robustness of transformer-based form field extraction methods via form attacks. We introduce 14 novel form transformations to evaluate…
InfoFocus: 3D Object Detection for Autonomous Driving with Dynamic Information Modeling
Jun Wang, Shiyi Lan, Mingfei Gao +1
Real-time 3D object detection is crucial for autonomous cars. Achieving promising performance with high efficiency, voxel-based approaches have received considerable attention. How…
WSLLN: Weakly Supervised Natural Language Localization Networks
Mingfei Gao, Larry S. Davis, Richard Socher +1
We propose weakly supervised language localization networks (WSLLN) to detect events in long, untrimmed videos given language queries. To learn the correspondence between visual se…
Goal-oriented Object Importance Estimation in On-road Driving Videos
Mingfei Gao, Ashish Tawari, Sujitha Martin
We formulate a new problem as Object Importance Estimation (OIE) in on-road driving videos, where the road users are considered as important objects if they have influence on the c…
StartNet: Online Detection of Action Start in Untrimmed Videos
Mingfei Gao, Mingze Xu, Larry S. Davis +2
We propose StartNet to address Online Detection of Action Start (ODAS) where action starts and their associated categories are detected in untrimmed, streaming videos. Previous met…