4 papers
Sidecar: Training-Free Semantic Reuse for Character-Consistent Free-form Visual Storytelling
Sibo Dong, Sarah Adel Bargal
Visual storytelling requires generating images that follow a narrative while preserving consistent character identities across frames. In free-form story generation, a character is…
FreeStory: Training-Free Character Consistency for Free-Form Visual Storytelling
Sibo Dong, Ismail Shaheen, Sarah Adel Bargal
Visual storytelling aims to generate image sequences that are both aligned with narrative prompts and consistent in character appearance across images. Recent training-free methods…
ViSTA: Visual Storytelling using Multi-modal Adapters for Text-to-Image Diffusion Models
Sibo Dong, Ismail Shaheen, Maggie Shen +2
Text-to-image diffusion models have achieved remarkable success, yet generating coherent image sequences for visual storytelling remains challenging. A key challenge is effectively…
D-Feat Occlusions: Diffusion Features for Robustness to Partial Visual Occlusions in Object Recognition
Rupayan Mallick, Sibo Dong, Nataniel Ruiz +1
Applications of diffusion models for visual tasks have been quite noteworthy. This paper targets making classification models more robust to occlusions for the task of object recog…