1 citations · 2 across the 6 of their papers we have counts for
8 papers
Motion-Guided Masking for Spatiotemporal Representation Learning
David Fan, Jue Wang, Shuai Liao +5
Several recent works have directly extended the image masked autoencoder (MAE) with random masking into video domain, achieving promising results. However, unlike images, both spat…
MEGA: Multimodal Alignment Aggregation and Distillation For Cinematic Video Segmentation
Najmeh Sadoughi, Xinyu Li, Avijit Vajpayee +5
Previous research has studied the task of segmenting cinematic videos into scenes and into narrative acts. However, these studies have overlooked the essential task of multimodal a…
Model-based Reconstruction for Multi-Frequency Collimated Beam Ultrasound Systems
Abdulrahman M. Alanazi, Singanallur Venkatakrishnan, Hector Santos-Villalobos +2
Collimated beam ultrasound systems are a technology for imaging inside multi-layered structures such as geothermal wells. These systems work by using a collimated narrow-band ultra…
Expanding Accurate Person Recognition to New Altitudes and Ranges: The BRIAR Dataset
David Cornett, Joel Brogan, Nell Barber +20
Face recognition technology has advanced significantly in recent years due largely to the availability of large and increasingly complex training datasets for use in deep learning…
Model-Based Reconstruction for Collimated Beam Ultrasound Systems
Abdulrahman Alanazi, Singanallur Venkatakrishnan, Hector Santos-Villalobos +2
Collimated beam ultrasound systems are a novel technology for imaging inside multi-layered structures such as geothermal wells. Such systems include a transmitter and multiple rece…
The Mertens Unrolled Network (MU-Net): A High Dynamic Range Fusion Neural Network for Through the Windshield Driver Recognition
Max Ruby, David S. Bolme, Joel Brogan +7
Face recognition of vehicle occupants through windshields in unconstrained environments poses a number of unique challenges ranging from glare, poor illumination, driver pose and m…