6 papers
EgoSelf: From Memory to Personalized Egocentric Assistant
Yanshuo Wang, Yuan Xu, Xuesong Li +4
Egocentric assistants often rely on first-person view data to capture user behavior and context for personalized services. Since different users exhibit distinct habits, preference…
Adaptive and Balanced Re-initialization for Long-timescale Continual Test-time Domain Adaptation
Yanshuo Wang, Jinguang Tong, Jun Lan +5
Continual test-time domain adaptation (CTTA) aims to adjust models so that they can perform well over time across non-stationary environments. While previous methods have made cons…
Structural Energy-Guided Sampling for View-Consistent Text-to-3D
Qing Zhang, Jinguang Tong, Jie Hong +2
Text-to-3D generation often suffers from the Janus problem, where objects look correct from the front but collapse into duplicated or distorted geometry from other angles. We attri…
Enhancing Features in Long-tailed Data Using Large Vision Model
Pengxiao Han, Changkun Ye, Jinguang Tong +4
Language-based foundation models, such as large language models (LLMs) or large vision-language models (LVLMs), have been widely studied in long-tailed recognition. However, the ne…
DGNS: Deformable Gaussian Splatting and Dynamic Neural Surface for Monocular Dynamic 3D Reconstruction
Xuesong Li, Jinguang Tong, Jie Hong +2
Dynamic scene reconstruction from monocular video is essential for real-world applications. We introduce DGNS, a hybrid framework integrating \underline{D}eformable \underline{G}au…
Improving Viewpoint Consistency in 3D Generation via Structure Feature and CLIP Guidance
Qing Zhang, Jinguang Tong, Jing Zhang +2
Despite recent advances in text-to-3D generation techniques, current methods often suffer from geometric inconsistencies, commonly referred to as the Janus Problem. This paper iden…