activity
20172024
most citedLocal-Global Context Aware Transformer for Language-Guided Video Segmentation

105 citations · 585 across the 48 of their papers we have counts for

collaborators
Showing 2021Show all

8 papers · 1 filter

cs.CV2021★ 18 cited

Collaborative Visual Navigation

Haiyang Wang, Wenguan Wang, Xizhou Zhu +2

As a fundamental problem for Artificial Intelligence, multi-agent system (MAS) is making rapid progress, mainly driven by multi-agent reinforcement learning (MARL) techniques. Howe…

cs.CV2021

A Survey on Deep Learning Technique for Video Segmentation

Tianfei Zhou, Fatih Porikli, David Crandall +2

Video segmentation -- partitioning video frames into multiple segments or objects -- plays a critical role in a broad range of practical applications, from enhancing visual effects…

cs.CV2021

Rethinking Cross-modal Interaction from a Top-down Perspective for Referring Video Object Segmentation

Chen Liang, Yu Wu, Tianfei Zhou +4

Referring video object segmentation (RVOS) aims to segment video objects with the guidance of natural language reference. Previous methods typically tackle RVOS through directly gr…

cs.CV2021

Collaborative Spatial-Temporal Modeling for Language-Queried Video Actor Segmentation

Tianrui Hui, Shaofei Huang, Si Liu +5

Language-queried video actor segmentation aims to predict the pixel-level mask of the actor which performs the actions described by a natural language query in the target frames. E…

cs.CV2021

Face Forensics in the Wild

Tianfei Zhou, Wenguan Wang, Zhiyuan Liang +1

On existing public benchmarks, face forgery detection techniques have achieved great success. However, when used in multi-person videos, which often contain many people active in t…

cs.CV2021★ 5 cited

Differentiable Multi-Granularity Human Representation Learning for Instance-Aware Human Semantic Parsing

Tianfei Zhou, Wenguan Wang, Si Liu +2

To address the challenging task of instance-aware human part parsing, a new bottom-up regime is proposed to learn category-level human semantic segmentation as well as multi-person…