7 papers
Catch Me If You Can Describe Me: Open-Vocabulary Camouflaged Instance Segmentation with Diffusion
Tuan-Anh Vu, Duc Thanh Nguyen, Qing Guo +4
Text-to-image diffusion techniques have shown exceptional capabilities in producing high-quality, dense visual predictions from open-vocabulary text. This indicates a strong correl…
Power of Boundary and Reflection: Semantic Transparent Object Segmentation using Pyramid Vision Transformer with Transparent Cues
Tuan-Anh Vu, Hai Nguyen-Truong, Ziqiang Zheng +4
Glass is a prevalent material among solid objects in everyday life, yet segmentation methods struggle to distinguish it from opaque materials due to its transparency and reflection…
MSC: A Marine Wildlife Video Dataset with Grounded Segmentation and Clip-Level Captioning
Quang-Trung Truong, Yuk-Kwan Wong, Vo Hoang Kim Tuyen Dang +3
Marine videos present significant challenges for video understanding due to the dynamics of marine objects and the surrounding environment, camera motion, and the complexity of und…
AUTV: Creating Underwater Video Datasets with Pixel-wise Annotations
Quang Trung Truong, Wong Yuk Kwan, Duc Thanh Nguyen +2
Underwater video analysis, hampered by the dynamic marine environment and camera motion, remains a challenging task in computer vision. Existing training-free video generation tech…
Color Alignment in Diffusion
Ka Chun Shum, Binh-Son Hua, Duc Thanh Nguyen +1
Diffusion models have shown great promise in synthesizing visually appealing images. However, it remains challenging to condition the synthesis at a fine-grained level, for instanc…
Advances in 3D Neural Stylization: A Survey
Yingshu Chen, Guocheng Shao, Ka Chun Shum +2
Modern artificial intelligence offers a novel and transformative approach to creating digital art across diverse styles and modalities like images, videos and 3D data, unleashing t…