4 papers
GeoDiff3D: Self-Supervised 3D Scene Generation with Geometry-Constrained 2D Diffusion Guidance
Haozhi Zhu, Miaomiao Zhao, Dingyao Liu +4
3D scene generation is a core technology for gaming, film/VFX, and VR/AR. Growing demand for rapid iteration, high-fidelity detail, and accessible content creation has further incr…
CombatVLA: An Efficient Vision-Language-Action Model for Combat Tasks in 3D Action Role-Playing Games
Peng Chen, Pi Bu, Yingyao Wang +9
Recent advances in Vision-Language-Action models (VLAs) have expanded the capabilities of embodied intelligence. However, significant challenges remain in real-time decision-making…
ArtCrafter: Text-Image Aligning Style Transfer via Embedding Reframing
Nisha Huang, Kaer Huang, Yifan Pu +5
Recent years have witnessed significant advancements in text-guided style transfer, primarily attributed to innovations in diffusion models. These models excel in conditional guida…
InterDance:Reactive 3D Dance Generation with Realistic Duet Interactions
Ronghui Li, Youliang Zhang, Yachao Zhang +6
Humans perform a variety of interactive motions, among which duet dance is one of the most challenging interactions. However, in terms of human motion generative models, existing w…