collaborators

10 papers

cs.CV2026

Compact Object-Level Representations with Open-Vocabulary Understanding for Indoor Visual Relocalization

Zhaopeng Cui, Jiarui Hu, Jingbo Liu +7

Indoor visual relocalization plays a critical role in emerging spatial and embodied AI applications. However, prior research was predominantly devoted to low-level vision schemes,…

cs.CV2026

NeuMesh++: Towards Versatile and Efficient Volumetric Editing with Disentangled Neural Mesh-based Implicit Field

Chong Bao, Yuan Li, Bangbang Yang +5

Recently neural implicit rendering techniques have evolved rapidly and demonstrated significant advantages in novel view synthesis and 3D scene reconstruction. However, existing ne…

cs.CV2026

Natural Human Motion Recovery by Aligning High-Order Temporal Dynamics from Monocular Videos

Dingkun Wei, Zehong Shen, Yan Xia +3

Human motion recovered from monocular videos often appears overly smooth or dynamically inconsistent, even when joint positions are numerically accurate. We observe that this limit…

cs.CV2026

DiffWind: Physics-Informed Differentiable Modeling of Wind-Driven Object Dynamics

Yuanhang Lei, Boming Zhao, Zesong Yang +10

Modeling wind-driven object dynamics from video observations is highly challenging due to the invisibility and spatio-temporal variability of wind, as well as the complex deformati…

cs.CV2025

BoxDreamer: Dreaming Box Corners for Generalizable Object Pose Estimation

Yuanhong Yu, Xingyi He, Chen Zhao +7

This paper presents a generalizable RGB-based approach for object pose estimation, specifically designed to address challenges in sparse-view settings. While existing methods can e…

cs.RO2025

Adaptive Visuo-Tactile Fusion with Predictive Force Attention for Dexterous Manipulation

Jinzhou Li, Tianhao Wu, Jiyao Zhang +6

Effectively utilizing multi-sensory data is important for robots to generalize across diverse tasks. However, the heterogeneous nature of these modalities makes fusion challenging.…