6 papers
Swift Sampling: Selecting Temporal Surprises via Taylor Series
Dahye Kim, Bhuvan Sachdeva, Karan Uppal +3
While most frames in long-form video are redundant, the critical information resides in temporal surprises: moments where the actual visual features deviate from their predicted ev…
Understanding Task Transfer in Vision-Language Models
Bhuvan Sachdeva, Karan Uppal, Abhinav Java +1
Vision-Language Models (VLMs) perform well on multimodal benchmarks but lag behind humans and specialized models on visual perception tasks like depth estimation or object counting…
CataractCompDetect: Intraoperative Complication Detection in Cataract Surgery
Bhuvan Sachdeva, Sneha Kumari, Rudransh Agarwal +8
Cataract surgery is one of the most commonly performed surgeries worldwide, yet intraoperative complications such as iris prolapse, posterior capsule rupture (PCR), and vitreous lo…
ASHABot: An LLM-Powered Chatbot to Support the Informational Needs of Community Health Workers
Pragnya Ramjee, Mehak Chhokar, Bhuvan Sachdeva +5
Community health workers (CHWs) provide last-mile healthcare services but face challenges due to limited medical knowledge and training. This paper describes the design, deployment…
CataractBot: An LLM-Powered Expert-in-the-Loop Chatbot for Cataract Patients
Pragnya Ramjee, Bhuvan Sachdeva, Satvik Golechha +4
The healthcare landscape is evolving, with patients seeking reliable information about their health conditions and available treatment options. Despite the abundance of information…
Phase-Informed Tool Segmentation for Manual Small-Incision Cataract Surgery
Bhuvan Sachdeva, Naren Akash, Tajamul Ashraf +6
Cataract surgery is the most common surgical procedure globally, with a disproportionately higher burden in developing countries. While automated surgical video analysis has been e…