3 papers
cs.CV2025
TB-HSU: Hierarchical 3D Scene Understanding with Contextual Affordances
Wenting Xu, Viorela Ila, Luping Zhou +1
The concept of function and affordance is a critical aspect of 3D scene understanding and supports task-oriented objectives. In this work, we develop a model that learns to structu…
cs.CV2024
Storynizor: Consistent Story Generation via Inter-Frame Synchronized and Shuffled ID Injection
Yuhang Ma, Wenting Xu, Chaoyi Zhao +5
Recent advances in text-to-image diffusion models have spurred significant interest in continuous story image generation. In this paper, we introduce Storynizor, a model capable of…
cs.CV2024
Character-Adapter: Prompt-Guided Region Control for High-Fidelity Character Customization
Yuhang Ma, Wenting Xu, Jiji Tang +5
Customized image generation, which seeks to synthesize images with consistent characters, holds significant relevance for applications such as storytelling, portrait generation, an…