41 citations · 99 across the 48 of their papers we have counts for
59 papers · 1 filter
Disparity Has a Sign: Stereo Matching Beyond the Zero-Disparity Plane
Jian Shi, Xinge Yang, Chaoyang Wang +2
Modern stereo matching models fail when disparity crosses zero, with end-point error (EPE) rising by 4.6-37. Yet stereoscopic content, from cinema 3D to VR, routinely conta…
EgoPlay: Event-Triggered Video Editing for Egocentric Streams
Jinjie Mai, Gordon Guocheng Qian, Willi Menapace +8
We introduce EgoPlay, an event-triggered video-to-video editor for egocentric streams, obtained by fine-tuning a pretrained V2V diffusion transformer on event-conditioned data buil…
RegHead: Non-Humanoid Head Blendshapes via Feed-Forward Registration
Jiahao Luo, Hao Zhang, Jianqi Chen +9
We present RegHead, a framework for constructing semantic blendshape sets for animatable non-humanoid head avatars. With a fixed expression vocabulary, semantic blendshapes provide…
MeshLoom: Feed-Forward Non-Rigid Registration of Mesh Sequences
Jianqi Chen, Jiraphon Yenphraphai, Xiangjun Tang +4
We present MeshLoom, a feed-forward registration network that directly reconstructs vertex deformations across mesh sequences. Our approach advances non-rigid registration beyond e…
GeoStream: Toward Precise Camera Controlled Streaming Video Generation
Yizhou Zhao, Yifan Wang, Xiaoyuan Wang +11
Accurate interactive camera control is essential for video-based world models, but most existing approaches learn camera motion implicitly, leading to inaccurate control under out-…
FROST-STA: Frozen Dense Features for the Ego4D Short-Term Object Interaction Anticipation
Chaoyang Wang, Lexuan Xu
Short-term anticipation in egocentric video requires more than recognizing the current scene: a system must infer which object the camera wearer will contact, which action will fol…