6 citations · 6 across the 3 of their papers we have counts for
4 papers
SceneTransporter: Optimal Transport-Guided Compositional Latent Diffusion for Single-Image Structured 3D Scene Generation
Ling Wang, Hao-Xiang Guo, Xinzhou Wang +9
We introduce SceneTransporter, an end-to-end framework for structured 3D scene generation from a single image. While existing methods generate part-level 3D objects, they often fai…
Video4DGen: Enhancing Video and 4D Generation through Mutual Optimization
Yikai Wang, Guangce Liu, Xinzhou Wang +5
The advancement of 4D (i.e., sequential 3D) generation opens up new possibilities for lifelike experiences in various applications, where users can explore dynamic objects or chara…
Medical Multimodal Foundation Models in Clinical Diagnosis and Treatment: Applications, Challenges, and Future Directions
Kai Sun, Siyan Xue, Fuchun Sun +12
Recent advancements in deep learning have significantly revolutionized the field of clinical diagnosis and treatment, offering novel approaches to improve diagnostic precision and…
Hunyuan3D 1.0: A Unified Framework for Text-to-3D and Image-to-3D Generation
Xianghui Yang, Huiwen Shi, Bowen Zhang +20
While 3D generative models have greatly improved artists' workflows, the existing diffusion models for 3D generation suffer from slow generation and poor generalization. To address…