activity
20242026
collaborators

9 papers

cs.CV2026

EchoStyle: Unlocking High-Fidelity Video Stylization with Reverse Data Synthesis

Huaqiu Li, Jiahao Wang, Sijia Cai +4

While image stylization has been studied extensively, video stylization remains a critical and largely unsolved challenge in the field of intelligent content creation. Existing met…

cs.CV2026

AnyID: Ultra-Fidelity Universal Identity-Preserving Video Generation from Any Visual References

Jiahao Wang, Hualian Sheng, Sijia Cai +5

Identity-preserving video generation offers powerful tools for creative expression, allowing users to customize videos featuring their beloved characters. However, prevailing metho…

cs.CV2026

DepthArb: Training-Free Depth-Arbitrated Generation for Occlusion-Robust Image Synthesis

Hongjin Niu, Jiahao Wang, Xirui Hu +4

Text-to-image models often struggle to synthesize correct occlusion relationships among multiple objects, especially in densely overlapping regions. Many training-free layout-guide…

cs.CV2025

DynamicID: Zero-Shot Multi-ID Image Personalization with Flexible Facial Editability

Xirui Hu, Jiahao Wang, Hao Chen +4

Recent advances in text-to-image generation have driven interest in generating personalized human images that depict specific identities from reference images. Although existing me…

cs.CV2025

EchoShot: Multi-Shot Portrait Video Generation

Jiahao Wang, Hualian Sheng, Sijia Cai +5

Video diffusion models substantially boost the productivity of artistic workflows with high-quality portrait video generative capacity. However, prevailing pipelines are primarily…

cs.CV2024

Accelerating Non-Maximum Suppression: A Graph Theory Perspective

King-Siong Si, Lu Sun, Weizhan Zhang +4

Non-maximum suppression (NMS) is an indispensable post-processing step in object detection. With the continuous optimization of network models, NMS has become the ``last mile'' to…