14 papers
FA-LAM: Focus-Aware Large Avatar Model for One-Shot 4D Animatable Gaussian Head
Yingdong Hu, Yisheng He, Yiming Jiang +3
We propose FA-LAM, a Focus-Aware Large Avatar Model for one-shot animatable Gaussian head creation, while simultaneously enabling static 3D and dynamic 4D full-head recovery. The c…
Bridging Visual and Wireless Sensing via a Unified Radiation Field for 3D Radio Map Construction
Chaozheng Wen, Jingwen Tong, Zehong Lin +2
The emerging applications of next-generation wireless networks demand high-fidelity environmental intelligence. 3D radio maps bridge physical environments and electromagnetic propa…
OCC: Physical-Layer Assisted Congestion Control for Real-Time Communications
Yufan Zhuang, Zili Meng, Zehong Lin +1
Real-time communications (RTC) is a core technology for emerging applications in 6G, such as cloud gaming, teleoperation, and extended reality (XR), which require consistently low…
Flash-VAED: Plug-and-Play VAE Decoders for Efficient Video Generation
Lunjie Zhu, Yushi Huang, Xingtong Ge +5
Latent diffusion models have enabled high-quality video synthesis, yet their inference remains costly and time-consuming. As diffusion transformers become increasingly efficient, t…
Mon3tr: Monocular 3D Telepresence with Pre-built Gaussian Avatars as Amortization
Fangyu Lin, Yingdong Hu, Zhening Liu +3
Immersive telepresence aims to transform human interaction in AR/VR applications by enabling lifelike full-body holographic representations for enhanced remote collaboration. Howev…
Feed-Forward 3D Gaussian Splatting Compression with Long-Context Modeling
Zhening Liu, Rui Song, Yushi Huang +5
3D Gaussian Splatting (3DGS) has emerged as a revolutionary 3D representation. However, its substantial data size poses a major barrier to widespread adoption. While feed-forward 3…