activity
20242026
collaborators

7 papers

cs.CV2026

Catch Me If You Can Describe Me: Open-Vocabulary Camouflaged Instance Segmentation with Diffusion

Tuan-Anh Vu, Duc Thanh Nguyen, Qing Guo +4

Text-to-image diffusion techniques have shown exceptional capabilities in producing high-quality, dense visual predictions from open-vocabulary text. This indicates a strong correl…

cs.CV2025

Power of Boundary and Reflection: Semantic Transparent Object Segmentation using Pyramid Vision Transformer with Transparent Cues

Tuan-Anh Vu, Hai Nguyen-Truong, Ziqiang Zheng +4

Glass is a prevalent material among solid objects in everyday life, yet segmentation methods struggle to distinguish it from opaque materials due to its transparency and reflection…

cs.CV2025

MSC: A Marine Wildlife Video Dataset with Grounded Segmentation and Clip-Level Captioning

Quang-Trung Truong, Yuk-Kwan Wong, Vo Hoang Kim Tuyen Dang +3

Marine videos present significant challenges for video understanding due to the dynamics of marine objects and the surrounding environment, camera motion, and the complexity of und…

cs.CE2025

AUTV: Creating Underwater Video Datasets with Pixel-wise Annotations

Quang Trung Truong, Wong Yuk Kwan, Duc Thanh Nguyen +2

Underwater video analysis, hampered by the dynamic marine environment and camera motion, remains a challenging task in computer vision. Existing training-free video generation tech…

cs.CV2025

Color Alignment in Diffusion

Ka Chun Shum, Binh-Son Hua, Duc Thanh Nguyen +1

Diffusion models have shown great promise in synthesizing visually appealing images. However, it remains challenging to condition the synthesis at a fine-grained level, for instanc…

cs.CV2024

Advances in 3D Neural Stylization: A Survey

Yingshu Chen, Guocheng Shao, Ka Chun Shum +2

Modern artificial intelligence offers a novel and transformative approach to creating digital art across diverse styles and modalities like images, videos and 3D data, unleashing t…