3 citations · 3 across the 3 of their papers we have counts for
3 papers
cs.CV2023
M3D: Learning 3D priors using Multi-Modal Masked Autoencoders for 2D image and video understanding
Muhammad Abdullah Jamal, Omid Mohareri
We present a new pre-training strategy called M3D (ulti-odal asked ) built based on Multi-modal masked autoencode…
cs.CV2023★ 3 cited
SurgMAE: Masked Autoencoders for Long Surgical Video Analysis
Muhammad Abdullah Jamal, Omid Mohareri
There has been a growing interest in using deep learning models for processing long surgical videos, in order to automatically detect clinical/operational activities and extract me…
cs.CV2022
Multi-Modal Unsupervised Pre-Training for Surgical Operating Room Workflow Analysis
Muhammad Abdullah Jamal, Omid Mohareri
Data-driven approaches to assist operating room (OR) workflow analysis depend on large curated datasets that are time consuming and expensive to collect. On the other hand, we see…