4 papers
4DLangVGGT: 4D Language-Visual Geometry Grounded Transformer
Xianfeng Wu, Yajing Bai, Minghan Li +5
Constructing 4D language fields is crucial for embodied AI, augmented/virtual reality, and 4D scene understanding, as they provide enriched semantic representations of dynamic envi…
LightGen: Efficient Image Generation through Knowledge Distillation and Direct Preference Optimization
Xianfeng Wu, Yajing Bai, Haoze Zheng +8
Recent advances in text-to-image generation have primarily relied on extensive datasets and parameter-heavy architectures. These requirements severely limit accessibility for resea…
Niagara: Normal-Integrated Geometric Affine Field for Scene Reconstruction from a Single View
Xianzu Wu, Zhenxin Ai, Harry Yang +3
Recent advances in single-view 3D scene reconstruction have highlighted the challenges in capturing fine geometric details and ensuring structural consistency, particularly in high…
Completing point cloud from few points by Wasserstein GAN and Transformers
Xianfeng Wu, Jinhui Qian, Qing Wei +6
In many vision and robotics applications, it is common that the captured objects are represented by very few points. Most of the existing completion methods are designed for partia…