activity
20242026
collaborators
Showing cs.CVShow all

7 papers · 1 filter

cs.CV2026

OmniColor: A Unified Framework for Multi-modal Lineart Colorization

Xulu Zhang, Haoqian Du, Xiaoyong Wei +1

Lineart colorization is a critical stage in professional content creation, yet achieving precise and flexible results under diverse user constraints remains a significant challenge…

cs.CV2025

A Survey on Personalized Content Synthesis with Diffusion Models

Xulu Zhang, Xiaoyong Wei, Wentao Hu +6

Recent advancements in diffusion models have significantly impacted content creation, leading to the emergence of Personalized Content Synthesis (PCS). By utilizing a small set of…

cs.CV2025

Mean of Means: Human Localization with Calibration-free and Unconstrained Camera Settings (extended version)

Tianyi Zhang, Wengyu Zhang, Xulu Zhang +4

Accurate human localization is crucial for various applications, especially in the Metaverse era. Existing high precision solutions rely on expensive, tag-dependent hardware, while…

cs.CV2025

Generating on Generated: An Approach Towards Self-Evolving Diffusion Models

Xulu Zhang, Xiaoyong Wei, Jinlin Wu +4

Recursive Self-Improvement (RSI) enables intelligence systems to autonomously refine their capabilities. This paper explores the application of RSI in text-to-image diffusion model…

cs.CV2025

Mean of Means: A 10-dollar Solution for Human Localization with Calibration-free and Unconstrained Camera Settings

Tianyi Zhang, Wengyu Zhang, Xulu Zhang +4

Accurate human localization is crucial for various applications, especially in the Metaverse era. Existing high precision solutions rely on expensive, tag-dependent hardware, while…

cs.CV2024

Prior Knowledge Integration via LLM Encoding and Pseudo Event Regulation for Video Moment Retrieval

Yiyang Jiang, Wengyu Zhang, Xulu Zhang +3

In this paper, we investigate the feasibility of leveraging large language models (LLMs) for integrating general knowledge and incorporating pseudo-events as priors for temporal co…