5 papers · 1 filter
PaceVGGT: Pre-Alternating-Attention Token Pruning for Visual Geometry Transformers
Haotang Li, Zhenyu Qi, Shaohan Henry Wang +5
Visual Geometry Transformer (VGGT) is a strong feed-forward model for multiple 3D tasks, but its Alternating-Attention (AA) stack scales quadratically in the total token count, mak…
MSSSeg: Learning Multi-Scale Structural Complexity for Self-Supervised Segmentation
Haotang Li, Zhenyu Qi, Hao Qin +4
Self-supervised semantic segmentation methods often suffer from structural errors, including merging distinct objects or fragmenting coherent regions, because they rely primarily o…
PhysDepth: Plug-and-Play Physical Refinement for Monocular Depth Estimation in Challenging Environments
Kebin Peng, Haotang Li, Zhenyu Qi +6
State-of-the-art monocular depth estimation (MDE) models often struggle in challenging environments, primarily because they overlook robust physical information. To demonstrate thi…
UWB-PostureGuard: A Privacy-Preserving RF Sensing System for Continuous Ergonomic Sitting Posture Monitoring
Haotang Li, Zhenyu Qi, Sen He +6
Improper sitting posture during prolonged computer use has become a significant public health concern. Traditional posture monitoring solutions face substantial barriers, including…
DynamicLip: Shape-Independent Continuous Authentication via Lip Articulator Dynamics
Huashan Chen, Yifan Xu, Yue Feng +6
Biometrics authentication has become increasingly popular due to its security and convenience; however, traditional biometrics are becoming less desirable in scenarios such as new…