activity
20212024
most citedSemantic Diffusion Network for Semantic Segmentation

13 citations · 26 across the 6 of their papers we have counts for

collaborators

6 papers

cs.CV2024

Ensemble Quadratic Assignment Network for Graph Matching

Haoru Tan, Chuang Wang, Sitong Wu +3

Graph matching is a commonly used technique in computer vision and pattern recognition. Recent data-driven approaches have improved the graph matching accuracy remarkably, whereas…

cs.LG20236 cited

Data Pruning via Moving-one-Sample-out

Haoru Tan, Sitong Wu, Fei Du +4

In this paper, we propose a novel data-pruning approach called moving-one-sample-out (MoSo), which aims to identify and remove the least informative samples from the training set.…

cs.CV20234 cited

RegionBLIP: A Unified Multi-modal Pre-training Framework for Holistic and Regional Comprehension

Qiang Zhou, Chaohui Yu, Shaofeng Zhang +3

In this work, we investigate extending the comprehension of Multi-modal Large Language Models (MLLMs) to regional objects. To this end, we propose to extract features corresponding…

cs.CV20233 cited

AxWin Transformer: A Context-Aware Vision Transformer Backbone with Axial Windows

Fangjian Lin, Yizhe Ma, Sitong Wu +2

Recently Transformer has shown good performance in several vision tasks due to its powerful modeling capabilities. To reduce the quadratic complexity caused by the attention, some…

cs.CV202313 cited

Semantic Diffusion Network for Semantic Segmentation

Haoru Tan, Sitong Wu, Jimin Pi

Precise and accurate predictions over boundary areas are essential for semantic segmentation. However, the commonly-used convolutional operators tend to smooth and blur local detai…

cs.CV2021

Pale Transformer: A General Vision Transformer Backbone with Pale-Shaped Attention

Sitong Wu, Tianyi Wu, Haoru Tan +1

Recently, Transformers have shown promising performance in various vision tasks. To reduce the quadratic computation complexity caused by the global self-attention, various methods…