2 citations · 3 across the 3 of their papers we have counts for
3 papers
Surgical-VQLA: Transformer with Gated Vision-Language Embedding for Visual Question Localized-Answering in Robotic Surgery
Long Bai, Mobarakol Islam, Lalithkumar Seenivasan +1
Despite the availability of computer-aided simulators and recorded videos of surgical procedures, junior residents still heavily rely on experts to answer their queries. However, e…
Paced-Curriculum Distillation with Prediction and Label Uncertainty for Image Segmentation
Mobarakol Islam, Lalithkumar Seenivasan, S. P. Sharan +4
Purpose: In curriculum learning, the idea is to train on easier samples first and gradually increase the difficulty, while in self-paced learning, a pacing function defines the spe…
Surgical-VQA: Visual Question Answering in Surgical Scenes using Transformer
Lalithkumar Seenivasan, Mobarakol Islam, Adithya K Krishna +1
Visual question answering (VQA) in surgery is largely unexplored. Expert surgeons are scarce and are often overloaded with clinical and academic workloads. This overload often limi…