Showing cs.CVShow all
2 papers · 1 filter
cs.CV2025
Surgical-LVLM: Learning to Adapt Large Vision-Language Model for Grounded Visual Question Answering in Robotic Surgery
Guankun Wang, Long Bai, Wan Jun Nah +7
Recent advancements in Surgical Visual Question Answering (Surgical-VQA) and related region grounding have shown great promise for robotic and medical applications, addressing the…
cs.CV2024
Text in the Dark: Extremely Low-Light Text Image Enhancement
Che-Tsung Lin, Chun Chet Ng, Zhi Qin Tan +7
Extremely low-light text images are common in natural scenes, making scene text detection and recognition challenging. One solution is to enhance these images using low-light image…