4 citations · 11 across the 22 of their papers we have counts for
16 papers
TaleForge: Interactive Multimodal System for Personalized Story Creation
Minh-Loi Nguyen, Quang-Khai Le, Tam V. Nguyen +2
Storytelling is a deeply personal and creative process, yet existing methods often treat users as passive consumers, offering generic plots with limited personalization. This under…
SAMURAI: Shape-Aware Multimodal Retrieval for 3D Object Identification
Dinh-Khoi Vo, Van-Loc Nguyen, Minh-Triet Tran +1
Retrieving 3D objects in complex indoor environments using only a masked 2D image and a natural language description presents significant challenges. The ROOMELSA challenge limits…
VisionGuard: Synergistic Framework for Helmet Violation Detection
Lam-Huy Nguyen, Thinh-Phuc Nguyen, Thanh-Hai Nguyen +3
Enforcing helmet regulations among motorcyclists is essential for enhancing road safety and ensuring the effectiveness of traffic management systems. However, automatic detection o…
Automated Image Recognition Framework
Quang-Binh Nguyen, Trong-Vu Hoang, Ngoc-Do Tran +3
While the efficacy of deep learning models heavily relies on data, gathering and annotating data for specific tasks, particularly when addressing novel or sensitive subjects lackin…
Efficient 3D Brain Tumor Segmentation with Axial-Coronal-Sagittal Embedding
Tuan-Luc Huynh, Thanh-Danh Le, Tam V. Nguyen +2
In this paper, we address the crucial task of brain tumor segmentation in medical imaging and propose innovative approaches to enhance its performance. The current state-of-the-art…
FaR: Enhancing Multi-Concept Text-to-Image Diffusion via Concept Fusion and Localized Refinement
Gia-Nghia Tran, Quang-Huy Che, Trong-Tai Dam Vu +4
Generating multiple new concepts remains a challenging problem in the text-to-image task. Current methods often overfit when trained on a small number of samples and struggle with…