2 papers
cs.CV2026
Rotation Equivariant Mamba for Vision Tasks
Zhongchen Zhao, Qi Xie, Keyu Huang +3
Rotation equivariance constitutes one of the most general and crucial structural priors for visual data, yet it remains notably absent from current Mamba-based vision architectures…
cs.CV2025
Polyline Path Masked Attention for Vision Transformer
Zhongchen Zhao, Chaodong Xiao, Hui Lin +3
Global dependency modeling and spatial position modeling are two core issues of the foundational architecture design in current deep learning frameworks. Recently, Vision Transform…