7 papers
MVHOI: Bridge Multi-view Condition to Complex Human-Object Interaction Video Reenactment via 3D Foundation Model
Jinguang Tong, Jinbo Wu, Kaisiyuan Wang +12
Human-Object Interaction (HOI) video reenactment aims to transfer the interaction dynamics of a source video to a novel target object while preserving realistic hand-object coordin…
Wound3DAssist: A Practical Framework for 3D Wound Assessment
Remi Chierchia, Rodrigo Santa Cruz, Léo Lebrat +7
Managing chronic wounds remains a major healthcare challenge, with clinical assessment often relying on subjective and time-consuming manual documentation methods. Although 2D digi…
DCHM: Depth-Consistent Human Modeling for Multiview Detection
Jiahao Ma, Tianyu Wang, Miaomiao Liu +2
Multiview pedestrian detection typically involves two stages: human modeling and pedestrian localization. Human modeling represents pedestrians in 3D space by fusing multiview info…
Efficient Depth- and Spatially-Varying Image Simulation for Defocus Deblur
Xinge Yang, Chuong Nguyen, Wenbin Wang +3
Modern cameras with large apertures often suffer from a shallow depth of field, resulting in blurry images of objects outside the focal plane. This limitation is particularly probl…
Puzzles: Unbounded Video-Depth Augmentation for Scalable End-to-End 3D Reconstruction
Jiahao Ma, Lei Wang, Miaomiao liu +2
Multi-view 3D reconstruction remains a core challenge in computer vision. Recent methods, such as DUST3R and its successors, directly regress pointmaps from image pairs without rel…
SOAF: Scene Occlusion-aware Neural Acoustic Field
Huiyu Gao, Jiahao Ma, David Ahmedt-Aristizabal +2
This paper tackles the problem of novel view audio-visual synthesis along an arbitrary trajectory in an indoor scene, given the audio-video recordings from other known trajectories…