9 citations · 14 across the 6 of their papers we have counts for
6 papers
Image Inpainting Models are Effective Tools for Instruction-guided Image Editing
Xuan Ju, Junhao Zhuang, Zhaoyang Zhang +3
This is the technique report for the winning solution of the CVPR2024 GenAI Media Generation Challenge Workshop's Instruction-guided Image Editing track. Instruction-guided image e…
MiraData: A Large-Scale Video Dataset with Long Durations and Structured Captions
Xuan Ju, Yiming Gao, Zhaoyang Zhang +6
Sora's high-motion intensity and long consistent videos have significantly impacted the field of video generation, attracting unprecedented attention. However, existing publicly av…
VLPose: Bridging the Domain Gap in Pose Estimation with Language-Vision Tuning
Jingyao Li, Pengguang Chen, Xuan Ju +2
Thanks to advances in deep learning techniques, Human Pose Estimation (HPE) has achieved significant progress in natural scenarios. However, these models perform poorly in artifici…
Direct Inversion: Boosting Diffusion-based Editing with 3 Lines of Code
Xuan Ju, Ailing Zeng, Yuxuan Bian +2
Text-guided diffusion models have revolutionized image generation and editing, offering exceptional realism and diversity. Specifically, in the context of diffusion-based editing,…
HumanSD: A Native Skeleton-Guided Diffusion Model for Human Image Generation
Xuan Ju, Ailing Zeng, Chenchen Zhao +3
Controllable human image generation (HIG) has numerous real-life applications. State-of-the-art solutions, such as ControlNet and T2I-Adapter, introduce an additional learnable bra…
Human-Art: A Versatile Human-Centric Dataset Bridging Natural and Artificial Scenes
Xuan Ju, Ailing Zeng, Jianan Wang +2
Humans have long been recorded in a variety of forms since antiquity. For example, sculptures and paintings were the primary media for depicting human beings before the invention o…