3 citations · 3 across the 4 of their papers we have counts for
5 papers
A Two-Stage Progressive Pre-training using Multi-Modal Contrastive Masked Autoencoders
Muhammad Abdullah Jamal, Omid Mohareri
In this paper, we propose a new progressive pre-training method for image understanding tasks which leverages RGB-D datasets. The method utilizes Multi-Modal Contrastive Masked Aut…
VidLPRO: A eo-anguage re-training Framework for botic and Laparoscopic Surgery
Mohammadmahdi Honarmand, Muhammad Abdullah Jamal, Omid Mohareri
We introduce VidLPRO, a novel video-language (VL) pre-training framework designed specifically for robotic and laparoscopic surgery. While existing surgical VL models primarily rel…
RGB-D Semantic SLAM for Surgical Robot Navigation in the Operating Room
Cong Gao, Dinesh Rabindran, Omid Mohareri
Gaining spatial awareness of the Operating Room (OR) for surgical robotic systems is a key technology that can enable intelligent applications aiming at improved OR workflow. In th…
Automatic Operating Room Surgical Activity Recognition for Robot-Assisted Surgery
Aidean Sharghi, Helene Haugerud, Daniel Oh +1
Automatic recognition of surgical activities in the operating room (OR) is a key technology for creating next generation intelligent surgical devices and workflow monitoring/suppor…
A Robotic 3D Perception System for Operating Room Environment Awareness
Zhaoshuo Li, Amirreza Shaban, Jean-Gabriel Simard +3
Purpose: We describe a 3D multi-view perception system for the da Vinci surgical system to enable Operating room (OR) scene understanding and context awareness. Methods: Our propos…