42 citations · 98 across the 21 of their papers we have counts for
3 papers · 2 filters
Generalized Product-of-Experts for Learning Multimodal Representations in Noisy Environments
Abhinav Joshi, Naman Gupta, Jinang Shah +3
A real-world application or setting involves interaction between different modalities (e.g., video, speech, text). In order to process the multimodal information automatically and…
Task-Aware Active Learning for Endoscopic Image Analysis
Shrawan Kumar Thapa, Pranav Poudel, Binod Bhattarai +1
Semantic segmentation of polyps and depth estimation are two important research problems in endoscopic image analysis. One of the main obstacles to conduct research on these resear…
Bimodal Camera Pose Prediction for Endoscopy
Anita Rau, Binod Bhattarai, Lourdes Agapito +1
Deducing the 3D structure of endoscopic scenes from images is exceedingly challenging. In addition to deformation and view-dependent lighting, tubular structures like the colon pre…