50 papers
Learning to Generate Multiple Objects from Dense and Occluded Layouts
Bach-Hoang Ngo, Si-Tri Ngo, Hieu Le +1
Text-to-image diffusion models fail to generate correct object counts in dense scenes, where overlapping instances collapse into indistinguishable structures despite appearing visu…
PhysMirror: Physics-Aware Mirror Object Generation
Xuan-Bach Mai, Duy-Phuc Nguyen, Quoc-Van Le +6
Synthesizing physically accurate mirror reflections remains a fundamental challenge for modern text-to-image diffusion models, which are increasingly critical for generating synthe…
AI-Generated Image Recognition via Fusion of CNNs and Vision Transformers
Xuan-Bach Mai, Hoang-Minh Nguyen-Huu, Quoc-Nghia Nguyen +3
Recent advancements in synthetic data technology have opened a new era where images of remarkable quality are generated, blurring the lines between real-life images and those produ…
Budget-Aware Keyboardless Interaction
Quang-Thang Nguyen, Gia-Phuc Song-Dong, Minh-Triet Tran +1
Interacting with computers typically relies on traditional input devices such as keyboards, mice, and monitors, which can be cumbersome for users seeking greater mobility. Virtual…
DanceDuo: Bridging Human Movement and AI Choreography
Gia-Cat Bui-Le, Tuong-Vy Truong-Thuy, Hai-Dang Nguyen +1
In recent years, advancements in deep learning and generative models have revolutionized music-driven dance generation. This paper introduces a novel platform, namely DanceDuo, lev…
KidRisk: Benchmark Dataset for Children Dangerous Action Recognition
Minh-Kha Nguyen, Trung-Hieu Do, Kim Anh Phung +3
Children are naturally energetic, and during their spontaneous activities, they often encounter potentially dangerous situations, especially when lacking parental supervision. Iden…