1 citations · 1 across the 1 of their papers we have counts for
2 papers
cs.CV2026
MarineEVT: Advancing Event-Centric Marine Video Understanding via Visual Tool Reasoning
Tuan-An To, Yuk-Kwan Wong, Tuan-Anh Vu +2
Recent Vision-Language Models (VLMs) have achieved remarkable success in visual understanding, driven by the growing availability of high-quality image-text pairs. However, the per…
cs.CV2025★ 1 cited
MarineEval: Assessing the Marine Intelligence of Vision-Language Models
YuK-Kwan Wong, Tuan-An To, Jipeng Zhang +2
We have witnessed promising progress led by large language models (LLMs) and further vision language models (VLMs) in handling various queries as a general-purpose assistant. VLMs,…