2 papers
cs.CV2025
MulCLIP: A Multi-level Alignment Framework for Enhancing Fine-grained Long-context CLIP
Chau Truong, Hieu Ta Quang, Dung D. Le
Vision-language models like CLIP show impressive ability to align images and text, but their training on short, concise captions makes them struggle with lengthy, detailed descript…
cs.CV2025
SDPA++: A General Framework for Self-Supervised Denoising with Patch Aggregation
Huy Minh Nhat Nguyen, Triet Hoang Minh Dao, Chau Vinh Hoang Truong +1
Optical Coherence Tomography (OCT) is a widely used non-invasive imaging technique that provides detailed three-dimensional views of the retina, which are essential for the early a…