5 papers
DreamVAR: Taming Reinforced Visual Autoregressive Model for High-Fidelity Subject-Driven Image Generation
Xin Jiang, Jingwen Chen, Yehao Li +5
Recent advances in subject-driven image generation using diffusion models have attracted considerable attention for their remarkable capabilities in producing high-quality images.…
IMAGGarment: Fine-Grained Garment Generation for Controllable Fashion Design
Fei Shen, Jian Yu, Cong Wang +3
This paper presents IMAGGarment, a fine-grained garment generation (FGG) framework that enables high-fidelity garment synthesis with precise control over silhouette, color, and log…
FaceShot: Bring Any Character into Life
Junyao Gao, Yanan Sun, Fei Shen +4
In this paper, we present FaceShot, a novel training-free portrait animation framework designed to bring any character into life from any driven video without fine-tuning or retrai…
@Bench: Benchmarking Vision-Language Models for Human-centered Assistive Technology
Xin Jiang, Junwei Zheng, Ruiping Liu +4
As Vision-Language Models (VLMs) advance, human-centered Assistive Technologies (ATs) for helping People with Visual Impairments (PVIs) are evolving into generalists, capable of pe…
NCST: Neural-based Color Style Transfer for Video Retouching
Xintao Jiang, Yaosen Chen, Siqin Zhang +2
Video color style transfer aims to transform the color style of an original video by using a reference style image. Most existing methods employ neural networks, which come with ch…