2 papers
cs.CV2025
Global-Local Aware Scene Text Editing
Fuxiang Yang, Tonghua Su, Donglin Di +4
Scene Text Editing (STE) involves replacing text in a scene image with new target text while preserving both the original text style and background texture. Existing methods suffer…
cs.CV2025
DiTalker: A Unified DiT-based Framework for High-Quality and Speaking Styles Controllable Portrait Animation
He Feng, Yongjia Ma, Donglin Di +3
Portrait animation aims to synthesize talking videos from a static reference face, conditioned on audio and style frame cues (e.g., emotion and head poses), while ensuring precise…