1 citations · 1 across the 6 of their papers we have counts for
9 papers
ESC: Emotional Self-Correction for Reliable Vision-Language Models
Tien-Huy Nguyen, Minh-Nhat Nguyen, Nguyen Nhat Huy +9
Vision-language models (VLMs) have achieved strong performance across diverse multimodal tasks, yet they remain vulnerable to unreliable reasoning. Existing self-correction methods…
ITSELF: Attention Guided Fine-Grained Alignment for Vision-Language Retrieval
Tien-Huy Nguyen, Huu-Loc Tran, Thanh Duc Ngo
Vision Language Models (VLMs) have rapidly advanced and show strong promise for text-based person search (TBPS), a task that requires capturing fine-grained relationships between i…
MADTempo: An Interactive System for Multi-Event Temporal Video Retrieval with Query Augmentation
Huu-An Vu, Van-Khanh Mai, Trong-Tam Nguyen +3
The rapid expansion of video content across online platforms has accelerated the need for retrieval systems capable of understanding not only isolated visual moments but also the t…
STER-VLM: Spatio-Temporal With Enhanced Reference Vision-Language Models
Tinh-Anh Nguyen-Nhu, Triet Dao Hoang Minh, Dat To-Thanh +3
Vision-language models (VLMs) have emerged as powerful tools for enabling automated traffic analysis; however, current approaches often demand substantial computational resources a…
IGL-DT: Iterative Global-Local Feature Learning with Dual-Teacher Semantic Segmentation Framework under Limited Annotation Scheme
Dinh Dai Quan Tran, Hoang-Thien Nguyen, Thanh-Huy Nguyen +3
Semi-Supervised Semantic Segmentation (SSSS) aims to improve segmentation accuracy by leveraging a small set of labeled images alongside a larger pool of unlabeled data. Recent adv…
HDC: Hierarchical Distillation for Multi-level Noisy Consistency in Semi-Supervised Fetal Ultrasound Segmentation
Tran Quoc Khanh Le, Nguyen Lan Vi Vu, Ha-Hieu Pham +5
Transvaginal ultrasound is a critical imaging modality for evaluating cervical anatomy and detecting physiological changes. However, accurate segmentation of cervical structures re…