activity
20242026
collaborators
Showing cs.CVShow all

6 papers · 1 filter

cs.CV2026

GeoCache: Training-Free Acceleration of Multi-View Texture Diffusion via Geometric Delta Transport

Haotang Li, Zhenyu Qi, Shaohan Henry Wang +6

Geometry-conditioned multi-view diffusion enables high-quality 3D texture generation, but its repeated per-view denoiser evaluations introduce substantial computational cost. Exist…

cs.CV2026

PaceVGGT: Pre-Alternating-Attention Token Pruning for Visual Geometry Transformers

Haotang Li, Zhenyu Qi, Shaohan Henry Wang +5

Visual Geometry Transformer (VGGT) is a strong feed-forward model for multiple 3D tasks, but its Alternating-Attention (AA) stack scales quadratically in the total token count, mak…

cs.CV2025

MSSSeg: Learning Multi-Scale Structural Complexity for Self-Supervised Segmentation

Haotang Li, Zhenyu Qi, Hao Qin +4

Self-supervised semantic segmentation methods often suffer from structural errors, including merging distinct objects or fragmenting coherent regions, because they rely primarily o…

cs.CV2025

UWB-PostureGuard: A Privacy-Preserving RF Sensing System for Continuous Ergonomic Sitting Posture Monitoring

Haotang Li, Zhenyu Qi, Sen He +6

Improper sitting posture during prolonged computer use has become a significant public health concern. Traditional posture monitoring solutions face substantial barriers, including…

cs.CV2025

DynamicLip: Shape-Independent Continuous Authentication via Lip Articulator Dynamics

Huashan Chen, Yifan Xu, Yue Feng +6

Biometrics authentication has become increasingly popular due to its security and convenience; however, traditional biometrics are becoming less desirable in scenarios such as new…

cs.CV2024

PhysDepth: Plug-and-Play Physical Refinement for Monocular Depth Estimation in Challenging Environments

Kebin Peng, Haotang Li, Zhenyu Qi +6

State-of-the-art monocular depth estimation (MDE) models often struggle in challenging environments, primarily because they overlook robust physical information. To demonstrate thi…