47 citations · 405 across the 64 of their papers we have counts for
70 papers · 1 filter
DexWorldModel: Causal Latent World Modeling towards Automated Learning of Embodied Tasks
Yueci Deng, Guiliang Liu, Kui Jia
Deploying generative World-Action Models for manipulation is severely bottlenecked by redundant pixel-level reconstruction, memory scaling, and sequential inferenc…
PAct: Part-Decomposed Single-View Articulated Object Generation
Qingming Liu, Xinyue Yao, Shuyuan Zhang +4
Articulated objects are central to interactive 3D applications, including embodied AI, robotics, and VR/AR, where functional part decomposition and kinematic motion are essential.…
Topology-Aware Modeling for Unsupervised Simulation-to-Reality Point Cloud Recognition
Longkun Zou, Kangjun Liu, Ke Chen +3
Learning semantic representations from point sets of 3D object shapes is often challenged by significant geometric variations, primarily due to differences in data acquisition meth…
SceneLCM: End-to-End Layout-Guided Interactive Indoor Scene Generation with Latent Consistency Model
Yangkai Lin, Jiabao Lei, Kui Jia
Our project page: https://scutyklin.github.io/SceneLCM/. Automated generation of complex, interactive indoor scenes tailored to user prompt remains a formidable challenge. While ex…
Multi-StyleGS: Stylizing Gaussian Splatting with Multiple Styles
Yangkai Lin, Jiabao Lei, Kui jia
In recent years, there has been a growing demand to stylize a given 3D scene to align with the artistic style of reference images for creative purposes. While 3D Gaussian Splatting…
Understanding Attention Mechanism in Video Diffusion Models
Bingyan Liu, Chengyu Wang, Tongtong Su +4
Text-to-video (T2V) synthesis models, such as OpenAI's Sora, have garnered significant attention due to their ability to generate high-quality videos from a text prompt. In diffusi…