2 papers
cs.CV2026
VisChronos: Revolutionizing Image Captioning Through Real-Life Events
Phuc-Tan Nguyen, Hieu Nguyen, Minh-Triet Tran +1
This paper aims to bridge the semantic gap between visual content and natural language understanding by leveraging historical events in the real world as a source of knowledge for…
cs.CV2025
OpenEvents V1: Large-Scale Benchmark Dataset for Multimodal Event Grounding
Hieu Nguyen, Phuc-Tan Nguyen, Thien-Phuc Tran +4
We introduce OpenEvents V1a large-scale benchmark dataset designed to advance event-centric vision-language understanding. Unlike conventional image captioning and retrieval datase…