7 papers
Rethinking Text-to-Image as Semantic-Aware Data Augmentation for Indoor Scene Recognition
Trong-Vu Hoang, Quang-Binh Nguyen, Dinh-Khoi Vo +3
In the realm of computer vision, indoor image recognition presents challenges due to the intricate interplay of lighting conditions, occlusions, and diverse object arrangements wit…
CSD-VAR: Content-Style Decomposition in Visual Autoregressive Models
Quang-Binh Nguyen, Minh Luu, Quang Nguyen +2
Disentangling content and style from a single image, known as content-style decomposition (CSD), enables recontextualization of extracted content and stylization of extracted style…
Automated Image Recognition Framework
Quang-Binh Nguyen, Trong-Vu Hoang, Ngoc-Do Tran +3
While the efficacy of deep learning models heavily relies on data, gathering and annotating data for specific tasks, particularly when addressing novel or sensitive subjects lackin…
ShowFlow: From Robust Single Concept to Condition-Free Multi-Concept Generation
Trong-Vu Hoang, Quang-Binh Nguyen, Thanh-Toan Do +3
Customizing image generation remains a core challenge in controllable image synthesis. For single-concept generation, maintaining both identity preservation and prompt alignment is…
ARtVista: Gateway To Empower Anyone Into Artist
Trong-Vu Hoang, Quang-Binh Nguyen, Duy-Nam Ly +4
Drawing is an art that enables people to express their imagination and emotions. However, individuals usually face challenges in drawing, especially when translating conceptual ide…
TextANIMAR: Text-based 3D Animal Fine-Grained Retrieval
Trung-Nghia Le, Tam V. Nguyen, Minh-Quan Le +30
3D object retrieval is an important yet challenging task that has drawn more and more attention in recent years. While existing approaches have made strides in addressing this issu…