most citedMSC: A Marine Wildlife Video Dataset with Grounded Segmentation and Clip-Level Captioning

1 citations · 1 across the 1 of their papers we have counts for

collaborators

5 papers

cs.CV20251 cited

MSC: A Marine Wildlife Video Dataset with Grounded Segmentation and Clip-Level Captioning

Quang-Trung Truong, Yuk-Kwan Wong, Vo Hoang Kim Tuyen Dang +3

Marine videos present significant challenges for video understanding due to the dynamics of marine objects and the surrounding environment, camera motion, and the complexity of und…

cs.CV2025

Adaptive Cache Enhancement for Test-Time Adaptation of Vision-Language Models

Khanh-Binh Nguyen, Phuoc-Nguyen Bui, Hyunseung Choo +1

Vision-language models (VLMs) exhibit remarkable zero-shot generalization but suffer performance degradation under distribution shifts in downstream tasks, particularly in the abse…

cs.CV2025

A model-agnostic active learning approach for animal detection from camera traps

Thi Thu Thuy Nguyen, Duc Thanh Nguyen

Smart data selection is becoming increasingly important in data-driven machine learning. Active learning offers a promising solution by allowing machine learning models to be effec…

cs.CE2025

AUTV: Creating Underwater Video Datasets with Pixel-wise Annotations

Quang Trung Truong, Wong Yuk Kwan, Duc Thanh Nguyen +2

Underwater video analysis, hampered by the dynamic marine environment and camera motion, remains a challenging task in computer vision. Existing training-free video generation tech…

cs.CV2025

Color Alignment in Diffusion

Ka Chun Shum, Binh-Son Hua, Duc Thanh Nguyen +1

Diffusion models have shown great promise in synthesizing visually appealing images. However, it remains challenging to condition the synthesis at a fine-grained level, for instanc…