collaborators
Showing cs.CVShow all

5 papers · 1 filter

cs.CV2026

PaceVGGT: Pre-Alternating-Attention Token Pruning for Visual Geometry Transformers

Haotang Li, Zhenyu Qi, Shaohan Henry Wang +5

Visual Geometry Transformer (VGGT) is a strong feed-forward model for multiple 3D tasks, but its Alternating-Attention (AA) stack scales quadratically in the total token count, mak…

cs.CV2026

MSSSeg: Learning Multi-Scale Structural Complexity for Self-Supervised Segmentation

Haotang Li, Zhenyu Qi, Hao Qin +4

Self-supervised semantic segmentation methods often suffer from structural errors, including merging distinct objects or fragmenting coherent regions, because they rely primarily o…

cs.CV2026

PhysDepth: Plug-and-Play Physical Refinement for Monocular Depth Estimation in Challenging Environments

Kebin Peng, Haotang Li, Zhenyu Qi +6

State-of-the-art monocular depth estimation (MDE) models often struggle in challenging environments, primarily because they overlook robust physical information. To demonstrate thi…

cs.CV2025

UWB-PostureGuard: A Privacy-Preserving RF Sensing System for Continuous Ergonomic Sitting Posture Monitoring

Haotang Li, Zhenyu Qi, Sen He +6

Improper sitting posture during prolonged computer use has become a significant public health concern. Traditional posture monitoring solutions face substantial barriers, including…

cs.CV2025

DynamicLip: Shape-Independent Continuous Authentication via Lip Articulator Dynamics

Huashan Chen, Yifan Xu, Yue Feng +6

Biometrics authentication has become increasingly popular due to its security and convenience; however, traditional biometrics are becoming less desirable in scenarios such as new…