1 citations · 1 across the 15 of their papers we have counts for
28 papers · 1 filter
Self-Supervised Multi-View 3D Gaze Target Estimation via Probabilistic Ray Marching
Keqi Chen, Vinkle Srivastav, Nicolas Padoy
We present a self-supervised approach, Self-MVGTE, for estimating 3D gaze targets from multiple camera views. Unlike existing methods that independently estimate 2D gaze targets pe…
SGRNet: Spatially Guided Radiology Network for Structured Radiological Reporting of Head and Neck Cancer
Ayush Gupta, Vinkle Srivastav, Prateek Upadhya +3
Automated radiological report generation can alleviate clinical workloads and eliminate observer variability. However, standard free-text generation models pose hallucination risks…
Where are they looking in the operating room?
Keqi Chen, Séraphin Baributsa, Lilien Schewski +5
Purpose: Gaze-following, the task of inferring where individuals are looking, has been widely studied in computer vision, advancing research in visual attention modeling, social sc…
CliPPER: Contextual Video-Language Pretraining on Long-form Intraoperative Surgical Procedures for Event Recognition
Florian Stilz, Vinkle Srivastav, Nassir Navab +1
Video-language foundation models have proven to be highly effective in zero-shot applications across a wide range of tasks. A particularly challenging area is the intraoperative su…
SurgTEMP: Temporal-Aware Surgical Video Question Answering with Text-guided Visual Memory for Laparoscopic Cholecystectomy
Shi Li, Vinkle Srivastav, Shih-Min Yin +7
Surgical procedures are inherently complex and risky, requiring extensive expertise and constant focus to navigate evolving intraoperative scenes. Computer-assisted systems such as…
Self-Supervised Uncalibrated Multi-View Video Anonymization in the Operating Room
Keqi Chen, Vinkle Srivastav, Armine Vardazaryan +3
Privacy preservation is a prerequisite for using video data in Operating Room (OR) research. Effective anonymization relies on the exhaustive localization of every individual; even…