48 citations · 53 across the 4 of their papers we have counts for
4 papers · 1 filter
BAVS: Bootstrapping Audio-Visual Segmentation by Integrating Foundation Knowledge
Chen Liu, Peike Li, Hu Zhang +4
Given an audio-visual pair, audio-visual segmentation (AVS) aims to locate sounding sources by predicting pixel-wise maps. Previous methods assume that each sound component in an a…
M6-Fashion: High-Fidelity Multi-modal Image Generation and Editing
Zhikang Li, Huiling Zhou, Shuai Bai +3
The fashion industry has diverse applications in multi-modal image generation and editing. It aims to create a desired high-fidelity image with the multi-modal conditional signal a…
Super-Resolving Cross-Domain Face Miniatures by Peeking at One-Shot Exemplar
Peike Li, Xin Yu, Yi Yang
Conventional face super-resolution methods usually assume testing low-resolution (LR) images lie in the same domain as the training ones. Due to different lighting conditions and i…
Self-Correction for Human Parsing
Peike Li, Yunqiu Xu, Yunchao Wei +1
Labeling pixel-level masks for fine-grained semantic segmentation tasks, e.g. human parsing, remains a challenging task. The ambiguous boundary between different semantic parts and…