15 citations · 22 across the 3 of their papers we have counts for
3 papers
cs.CV2023★ 3 cited
AxWin Transformer: A Context-Aware Vision Transformer Backbone with Axial Windows
Fangjian Lin, Yizhe Ma, Sitong Wu +2
Recently Transformer has shown good performance in several vision tasks due to its powerful modeling capabilities. To reduce the quadratic complexity caused by the attention, some…
cs.CV2023★ 4 cited
Exploring vision transformer layer choosing for semantic segmentation
Fangjian Lin, Yizhe Ma, Shengwei Tian
Extensive work has demonstrated the effectiveness of Vision Transformers. The plain Vision Transformer tends to obtain multi-scale features by selecting fixed layers, or the last l…
cs.CV2023★ 15 cited
PRSeg: A Lightweight Patch Rotate MLP Decoder for Semantic Segmentation
Yizhe Ma, Fangjian Lin, Sitong Wu +2
The lightweight MLP-based decoder has become increasingly promising for semantic segmentation. However, the channel-wise MLP cannot expand the receptive fields, lacking the context…