7 papers · 1 filter
OmniColor: A Unified Framework for Multi-modal Lineart Colorization
Xulu Zhang, Haoqian Du, Xiaoyong Wei +1
Lineart colorization is a critical stage in professional content creation, yet achieving precise and flexible results under diverse user constraints remains a significant challenge…
A Survey on Personalized Content Synthesis with Diffusion Models
Xulu Zhang, Xiaoyong Wei, Wentao Hu +6
Recent advancements in diffusion models have significantly impacted content creation, leading to the emergence of Personalized Content Synthesis (PCS). By utilizing a small set of…
Mean of Means: Human Localization with Calibration-free and Unconstrained Camera Settings (extended version)
Tianyi Zhang, Wengyu Zhang, Xulu Zhang +4
Accurate human localization is crucial for various applications, especially in the Metaverse era. Existing high precision solutions rely on expensive, tag-dependent hardware, while…
Generating on Generated: An Approach Towards Self-Evolving Diffusion Models
Xulu Zhang, Xiaoyong Wei, Jinlin Wu +4
Recursive Self-Improvement (RSI) enables intelligence systems to autonomously refine their capabilities. This paper explores the application of RSI in text-to-image diffusion model…
Mean of Means: A 10-dollar Solution for Human Localization with Calibration-free and Unconstrained Camera Settings
Tianyi Zhang, Wengyu Zhang, Xulu Zhang +4
Accurate human localization is crucial for various applications, especially in the Metaverse era. Existing high precision solutions rely on expensive, tag-dependent hardware, while…
Prior Knowledge Integration via LLM Encoding and Pseudo Event Regulation for Video Moment Retrieval
Yiyang Jiang, Wengyu Zhang, Xulu Zhang +3
In this paper, we investigate the feasibility of leveraging large language models (LLMs) for integrating general knowledge and incorporating pseudo-events as priors for temporal co…