12 citations · 17 across the 4 of their papers we have counts for
4 papers · 1 filter
OmniEdit: A Training-free framework for Lip Synchronization and Audio-Visual Editing
Lixiang Lin, Siyuan Jin, Jinshan Zhang
Lip synchronization and audio-visual editing have emerged as fundamental challenges in multimodal learning, underpinning a wide range of applications, including film production, vi…
Semantic-Preserved Point-based Human Avatar
Lixiang Lin, Jianke Zhu
To enable realistic experience in AR/VR and digital entertainment, we present the first point-based human avatar model that embodies the entirety expressive range of digital humans…
Weakly-Supervised Multi-Face 3D Reconstruction
Jialiang Zhang, Lixiang Lin, Jianke Zhu +1
3D face reconstruction plays a very important role in many real-world multimedia applications, including digital entertainment, social media, affection analysis, and person identif…
Attribute-aware Pedestrian Detection in a Crowd
Jialiang Zhang, Lixiang Lin, Yang Li +4
Pedestrian detection is an initial step to perform outdoor scene analysis, which plays an essential role in many real-world applications. Although having enjoyed the merits of deep…