papers

Publications (29)

cs.CV2022

Conditional Hyper-Network for Blind Super-Resolution with Multiple Degradations

Guanghao Yin, Wei Wang, Zehuan Yuan +5

Although single-image super-resolution (SISR) methods have achieved great success on single degradation, they still suffer performance drop with multiple degrading effects in real…

cs.CV2022

ByteTrack: Multi-Object Tracking by Associating Every Detection Box

Yifu Zhang, Peize Sun, Yi Jiang +6

Multi-object tracking (MOT) aims at estimating bounding boxes and identities of objects in videos. Most methods obtain identities by associating detection boxes whose scores are hi…

cs.CV2025

MMGen: Unified Multi-modal Image Generation and Understanding in One Go

Jiepeng Wang, Zhaoqing Wang, Hao Pan +4

A unified diffusion framework for multi-modal generation and understanding has the transformative potential to achieve seamless and controllable image diffusion and other cross-mod…

cs.CV2022

Single-Stage Open-world Instance Segmentation with Cross-task Consistency Regularization

Xizhe Xue, Dongdong Yu, Lingqiao Liu +6

Open-World Instance Segmentation (OWIS) is an emerging research topic that aims to segment class-agnostic object instances from images. The mainstream approaches use a two-stage se…

cs.CV2019

Towards Good Practices for Video Object Segmentation

Dongdong Yu, Kai Su, Hengkai Guo +6

Semi-supervised video object segmentation is an interesting yet challenging task in machine learning. In this work, we conduct a series of refinements with the propagation-based vi…

cs.CV2021

Mining Contextual Information Beyond Image for Semantic Segmentation

Zhenchao Jin, Tao Gong, Dongdong Yu +4

This paper studies the context aggregation problem in semantic image segmentation. The existing researches focus on improving the pixel representations by aggregating the contextua…