7 papers
Real2Sim2Real for Vision-Language-Action Manipulation: An AMD ROCm-Based Pipeline
Qing Yang, Xun Wang, Ziguan Wang +3
Physical AI -- the integration of large vision-language-action (VLA) models with embodied agents that act in the real world -- has emerged as the next major frontier for AI, echoed…
TagSplat: Topology-Aware Gaussian Splatting for Dynamic Mesh Modeling and Tracking
Hanzhi Guo, Dongdong Weng, Mo Su +3
Topology-consistent dynamic model sequences are essential for applications such as animation and model editing. However, existing 4D reconstruction methods face challenges in gener…
LoDAvatar: Hierarchical Embedding and Selective Detail Enhancement for Adaptive Levels of Detail Gaussian Avatars
Xiaonuo Dongye, Hanzhi Guo, Le Luo +5
With the advancement of virtual reality, the demand for 3D human avatars is increasing. The emergence of Gaussian Splatting technology has enabled the rendering of Gaussian avatars…
PS-GS: Gaussian Splatting for Multi-View Photometric Stereo
Yixiao Chen, Bin Liang, Hanzhi Guo +3
Integrating inverse rendering with multi-view photometric stereo (MVPS) yields more accurate 3D reconstructions than the inverse rendering approaches that rely on fixed environment…
STGA: Selective-Training Gaussian Head Avatars
Hanzhi Guo, Yixiao Chen, Dongye Xiaonuo +3
We propose selective-training Gaussian head avatars (STGA) to enhance the details of dynamic head Gaussian. The dynamic head Gaussian model is trained based on the FLAME parameteri…
Motion Generation Review: Exploring Deep Learning for Lifelike Animation with Manifold
Jiayi Zhao, Dongdong Weng, Qiuxin Du +1
Human motion generation involves creating natural sequences of human body poses, widely used in gaming, virtual reality, and human-computer interaction. It aims to produce lifelike…