2 citations · 4 across the 3 of their papers we have counts for
3 papers
cs.CV2023
ViLTA: Enhancing Vision-Language Pre-training through Textual Augmentation
Weihan Wang, Zhen Yang, Bin Xu +2
Vision-language pre-training (VLP) methods are blossoming recently, and its crucial goal is to jointly learn visual and textual features via a transformer-based architecture, demon…
cs.CV2012★ 2 cited
A 3D Segmentation Method for Retinal Optical Coherence Tomography Volume Data
Yankui Sun, Tian Zhang
With the introduction of spectral-domain optical coherence tomography (OCT), much larger image datasets are routinely acquired compared to what was possible using the previous gene…
cs.CV2012★ 2 cited
OCT Segmentation Survey and Summary Reviews and a Novel 3D Segmentation Algorithm and a Proof of Concept Implementation
Serguei A. Mokhov, Yankui Sun
We overview the existing OCT work, especially the practical aspects of it. We create a novel algorithm for 3D OCT segmentation with the goals of speed and/or accuracy while remaini…