activity
20182024
most citedTorchreid: A Library for Deep Learning Person Re-Identification in Pytorch

99 citations · 291 across the 9 of their papers we have counts for

collaborators
Showing cs.CVShow all

17 papers · 1 filter

cs.CV2024

4D Panoptic Scene Graph Generation

Jingkang Yang, Jun Cen, Wenxuan Peng +6

We are living in a three-dimensional space while moving forward through a fourth dimension: time. To allow artificial intelligence to develop a comprehensive understanding of such…

cs.CV2023

Panoptic Video Scene Graph Generation

Jingkang Yang, Wenxuan Peng, Xiangtai Li +8

Towards building comprehensive real-world visual perception systems, we propose and study a new problem called panoptic scene graph generation (PVSG). PVSG relates to the existing…

cs.CV2023

Octopus: Embodied Vision-Language Programmer from Environmental Feedback

Jingkang Yang, Yuhao Dong, Shuai Liu +8

Large vision-language models (VLMs) have achieved substantial progress in multimodal perception and reasoning. When integrated into an embodied agent, existing embodied VLM works e…

cs.CV2023

Semi-Supervised and Long-Tailed Object Detection with CascadeMatch

Yuhang Zang, Kaiyang Zhou, Chen Huang +1

This paper focuses on long-tailed object detection in the semi-supervised learning setting, which poses realistic challenges, but has rarely been studied in the literature. We prop…

cs.CV2023

Contextual Object Detection with Multimodal Large Language Models

Yuhang Zang, Wei Li, Jun Han +2

Recent Multimodal Large Language Models (MLLMs) are remarkable in vision-language tasks, such as image captioning and question answering, but lack the essential perception ability,…

cs.CV2022

Learning to Augment via Implicit Differentiation for Domain Generalization

Tingwei Wang, Da Li, Kaiyang Zhou +2

Machine learning models are intrinsically vulnerable to domain shift between training and testing data, resulting in poor performance in novel domains. Domain generalization (DG) a…