6 papers
O-DisCo-Edit: Object Distortion Control for Unified Realistic Video Editing
Yuqing Chen, Junjie Wang, Lin Liu +4
Diffusion models have recently advanced video editing, yet controllable editing remains challenging due to the need for precise manipulation of diverse object properties. Current m…
Bind-Your-Avatar: Multi-Talking-Character Video Generation with Dynamic 3D-mask-based Embedding Router
Yubo Huang, Weiqiang Wang, Sirui Zhao +3
Recent years have witnessed remarkable advances in audio-driven talking head generation. However, existing approaches predominantly focus on single-character scenarios. While some…
S2R-Bench: A Sim-to-Real Evaluation Benchmark for Autonomous Driving
Li Wang, Guangqi Yang, Lei Yang +13
Safety is a long-standing and the final pursuit in the development of autonomous driving systems, with a significant portion of safety challenge arising from perception. How to eff…
A Novel Generative Model with Causality Constraint for Mitigating Biases in Recommender Systems
Jianfeng Deng, Qingfeng Chen, Debo Cheng +3
Accurately predicting counterfactual user feedback is essential for building effective recommender systems. However, latent confounding bias can obscure the true causal relationshi…
UNSCT-HRNet: Modeling Anatomical Uncertainty for Landmark Detection in Total Hip Arthroplasty
Jiaxin Wan, Lin Liu, Haoran Wang +9
Total hip arthroplasty (THA) relies on accurate landmark detection from radiographic images, but unstructured data caused by irregular patient postures or occluded anatomical marke…
Three-Dimensional Medical Image Fusion with Deformable Cross-Attention
Lin Liu, Xinxin Fan, Chulong Zhang +3
Multimodal medical image fusion plays an instrumental role in several areas of medical image processing, particularly in disease recognition and tumor detection. Traditional fusion…