activity
20202022
most citedGRiT: A Generative Region-to-text Transformer for Object Understanding

30 citations · 46 across the 4 of their papers we have counts for

collaborators

5 papers

cs.CV202230 cited

GRiT: A Generative Region-to-text Transformer for Object Understanding

Jialian Wu, Jianfeng Wang, Zhengyuan Yang +4

This paper presents a Generative RegIon-to-Text transformer, GRiT, for object understanding. The spirit of GRiT is to formulate object understanding as <region, text> pairs, where…

cs.CV2022

Deformable VisTR: Spatio temporal deformable attention for video instance segmentation

Sudhir Yarram, Jialian Wu, Pan Ji +2

Video instance segmentation (VIS) task requires classifying, segmenting, and tracking object instances over all frames in a video clip. Recently, VisTR has been proposed as end-to-…

cs.CV20224 cited

Efficient Video Instance Segmentation via Tracklet Query and Proposal

Jialian Wu, Sudhir Yarram, Hui Liang +4

Video Instance Segmentation (VIS) aims to simultaneously classify, segment, and track multiple object instances in videos. Recent clip-level VIS takes a short video clip as input e…

cs.CV202112 cited

Track to Detect and Segment: An Online Multi-Object Tracker

Jialian Wu, Jiale Cao, Liangchen Song +3

Most online multi-object trackers perform object detection stand-alone in a neural net without any input from tracking. In this paper, we present a new online joint detection and t…

cs.CV2020

Forest R-CNN: Large-Vocabulary Long-Tailed Object Detection and Instance Segmentation

Jialian Wu, Liangchen Song, Tiancai Wang +2

Despite the previous success of object analysis, detecting and segmenting a large number of object categories with a long-tailed data distribution remains a challenging problem and…