activity
20172023
most citedDeeply-Learned Part-Aligned Representations for Person Re-Identification

166 citations · 536 across the 55 of their papers we have counts for

collaborators

59 papers

cs.CV2023★ 5 cited

Dancing Avatar: Pose and Text-Guided Human Motion Videos Synthesis with Image Diffusion Model

Bosheng Qin, Wentao Ye, Qifan Yu +2

The rising demand for creating lifelike avatars in the digital realm has led to an increased need for generating high-quality human videos guided by textual descriptions and poses.…

cs.CV2023

Degeneration-Tuning: Using Scrambled Grid shield Unwanted Concepts from Stable Diffusion

Zixuan Ni, Longhui Wei, Jiacheng Li +3

Owing to the unrestricted nature of the content in the training data, large text-to-image diffusion models, such as Stable Diffusion (SD), are capable of generating images with pot…

cs.CV2023★ 11 cited

Fine-tuning Multimodal LLMs to Follow Zero-shot Demonstrative Instructions

Juncheng Li, Kaihang Pan, Zhiqi Ge +7

Recent advancements in Multimodal Large Language Models (MLLMs) have been utilizing Visual Prompt Generators (VPGs) to convert visual features into tokens that LLMs can recognize.…

cs.CV2023

ZJU ReLER Submission for EPIC-KITCHEN Challenge 2023: TREK-150 Single Object Tracking

Yuanyou Xu, Jiahao Li, Zongxin Yang +2

The Associating Objects with Transformers (AOT) framework has exhibited exceptional performance in a wide range of complex scenarios for video object tracking and segmentation. In…

cs.CV2023

ZJU ReLER Submission for EPIC-KITCHEN Challenge 2023: Semi-Supervised Video Object Segmentation

Jiahao Li, Yuanyou Xu, Zongxin Yang +2

The Associating Objects with Transformers (AOT) framework has exhibited exceptional performance in a wide range of complex scenarios for video object segmentation. In this study, w…

cs.CV2023★ 2 cited

Improving Reference-based Distinctive Image Captioning with Contrastive Rewards

Yangjun Mao, Jun Xiao, Dong Zhang +4

Distinctive Image Captioning (DIC) -- generating distinctive captions that describe the unique details of a target image -- has received considerable attention over the last few ye…