3 papers
cs.CV2025
MT-Mark: Rethinking Image Watermarking via Mutual-Teacher Collaboration with Adaptive Feature Modulation
Fei Ge, Ying Huang, Jie Liu +4
Existing deep image watermarking methods follow a fixed embedding-distortion-extraction pipeline, where the embedder and extractor are weakly coupled through a final loss and optim…
cs.HC2024
A Unified Editing Method for Co-Speech Gesture Generation via Diffusion Inversion
Zeyu Zhao, Nan Gao, Zhi Zeng +3
Diffusion models have shown great success in generating high-quality co-speech gestures for interactive humanoid robots or digital avatars from noisy input with the speech audio or…
cs.CV2024
DiffMAC: Diffusion Manifold Hallucination Correction for High Generalization Blind Face Restoration
Nan Gao, Jia Li, Huaibo Huang +4
Blind face restoration (BFR) is a highly challenging problem due to the uncertainty of degradation patterns. Current methods have low generalization across photorealistic and heter…