Showing cs.CVShow all
3 papers · 1 filter
cs.CV2025
Zero-shot Reconstruction of In-Scene Object Manipulation from Video
Dixuan Lin, Tianyou Wang, Zhuoyang Pan +3
We build the first system to address the problem of reconstructing in-scene object manipulation from a monocular RGB video. It is challenging due to ill-posed scene reconstruction,…
cs.CV2024
OmniHands: Towards Robust 4D Hand Mesh Recovery via A Versatile Transformer
Dixuan Lin, Yuxiang Zhang, Mengcheng Li +5
In this paper, we introduce OmniHands, a universal approach to recovering interactive hand meshes and their relative movement from monocular or multi-view inputs. Our approach addr…
cs.CV2023
Cross-Modal Adaptive Dual Association for Text-to-Image Person Retrieval
Dixuan Lin, Yixing Peng, Jingke Meng +1
Text-to-image person re-identification (ReID) aims to retrieve images of a person based on a given textual description. The key challenge is to learn the relations between detailed…