Showing cs.CVShow all
3 papers · 1 filter
cs.CV2026
An Ultra-Widefield Swept-Source OCTA Dataset and a Polar-Gated Mamba Network for Retinal Vessel Segmentation
Yang Liu, Yibing Shen, Keming Zhao +12
Ultra-widefield (UWF) swept-source optical coherence tomography angiography (SS-OCTA) enables large-area retinal vascular imaging, yet vessel segmentation at this scale lacks dedic…
cs.CV2025
Hear-Your-Click: Interactive Object-Specific Video-to-Audio Generation
Yingshan Liang, Keyu Fan, Zhicheng Du +5
Video-to-audio (V2A) generation shows great potential in fields such as film production. Despite significant advances, current V2A methods relying on global video information strug…
cs.CV2024
Cognitive resilience: Unraveling the proficiency of image-captioning models to interpret masked visual content
Zhicheng Du, Zhaotian Xie, Huazhang Ying +2
This study explores the ability of Image Captioning (IC) models to decode masked visual content sourced from diverse datasets. Our findings reveal the IC model's capability to gene…