4.4k citations · 4.9k across the 45 of their papers we have counts for
9 papers · 2 filters
3DMODT: Attention-Guided Affinities for Joint Detection & Tracking in 3D Point Clouds
Jyoti Kini, Ajmal Mian, Mubarak Shah
We propose a method for joint detection and tracking of multiple objects in 3D point clouds, a task conventionally treated as a two-step process comprising object detection followe…
TransVisDrone: Spatio-Temporal Transformer for Vision-based Drone-to-Drone Detection in Aerial Videos
Tushar Sangam, Ishan Rajendrakumar Dave, Waqas Sultani +1
Drone-to-drone detection using visual feed has crucial applications, such as detecting drone collisions, detecting drone attacks, or coordinating flight with other drones. However,…
GAMa: Cross-view Video Geo-localization
Shruti Vyas, Chen Chen, Mubarak Shah
The existing work in cross-view geo-localization is based on images where a ground panorama is matched to an aerial image. In this work, we focus on ground videos instead of images…
Self-Supervised Video Object Segmentation via Cutout Prediction and Tagging
Jyoti Kini, Fahad Shahbaz Khan, Salman Khan +1
We propose a novel self-supervised Video Object Segmentation (VOS) approach that strives to achieve better object-background discriminability for accurate object segmentation. Dist…
Tag-Based Attention Guided Bottom-Up Approach for Video Instance Segmentation
Jyoti Kini, Mubarak Shah
Video Instance Segmentation is a fundamental computer vision task that deals with segmenting and tracking object instances across a video sequence. Most existing methods typically…
Video Action Detection: Analysing Limitations and Challenges
Rajat Modi, Aayush Jung Rana, Akash Kumar +4
Beyond possessing large enough size to feed data hungry machines (eg, transformers), what attributes measure the quality of a dataset? Assuming that the definitions of such attribu…