Showing cs.CVShow all
3 papers · 1 filter
cs.CV2024
ConsistentAvatar: Learning to Diffuse Fully Consistent Talking Head Avatar with Temporal Guidance
Haijie Yang, Zhenyu Zhang, Hao Tang +2
Diffusion models have shown impressive potential on talking head generation. While plausible appearance and talking effect are achieved, these methods still suffer from temporal, 3…
cs.CV2024
StoryImager: A Unified and Efficient Framework for Coherent Story Visualization and Completion
Ming Tao, Bing-Kun Bao, Hao Tang +2
Story visualization aims to generate a series of realistic and coherent images based on a storyline. Current models adopt a frame-by-frame architecture by transforming the pre-trai…
cs.CV2023
Adapting Segment Anything Model for Change Detection in HR Remote Sensing Images
Lei Ding, Kun Zhu, Daifeng Peng +3
Vision Foundation Models (VFMs) such as the Segment Anything Model (SAM) allow zero-shot or interactive segmentation of visual contents, thus they are quickly applied in a variety…