papers

Publications (30)

cs.CV2020

Multi-Drone based Single Object Tracking with Agent Sharing Network

Pengfei Zhu, Jiayu Zheng, Dawei Du +3

Drone equipped with cameras can dynamically track the target in the air from a broader view compared with static cameras or moving sensors over the ground. However, it is still cha…

cs.CV2024

Multi-view Deep Subspace Clustering Networks

Pengfei Zhu, Xinjie Yao, Yu Wang +3

Multi-view subspace clustering aims to discover the inherent structure of data by fusing multiple views of complementary information. Most existing methods first extract multiple t…

cs.CV2019

Scale Invariant Fully Convolutional Network: Detecting Hands Efficiently

Dan Liu, Dawei Du, Libo Zhang +4

Existing hand detection methods usually follow the pipeline of multiple stages with high computation cost, i.e., feature extraction, region proposal, bounding box regression, and a…

cs.CV2026

VEBench:Benchmarking Large Multimodal Models for Real-World Video Editing

Andong Deng, Dawei Du, Zhenfang Chen +7

Real-world video editing demands not only expert knowledge of cinematic techniques but also multimodal reasoning to select, align, and combine footage into coherent narratives. Whi…

cs.CV2022

Cascade Transformers for End-to-End Person Search

Rui Yu, Dawei Du, Rodney LaLonde +4

The goal of person search is to localize a target person from a gallery set of scene images, which is extremely challenging due to large scale variations, pose/viewpoint changes, a…

cs.CV2019

Learning Semantic Neural Tree for Human Parsing

Ruyi Ji, Dawei Du, Libo Zhang +5

The majority of existing human parsing methods formulate the task as semantic segmentation, which regard each semantic category equally and fail to exploit the intrinsic physiologi…