Showing cs.CVShow all
3 papers · 1 filter
cs.CV2025
Brain-Inspired Stepwise Patch Merging for Vision Transformers
Yonghao Yu, Dongcheng Zhao, Guobin Shen +2
The hierarchical architecture has become a mainstream design paradigm for Vision Transformers (ViTs), with Patch Merging serving as the pivotal component that transforms a columnar…
cs.CV2025
Enhancing Audio-Visual Spiking Neural Networks through Semantic-Alignment and Cross-Modal Residual Learning
Xiang He, Dongcheng Zhao, Yiting Dong +3
Humans interpret and perceive the world by integrating sensory information from multiple modalities, such as vision and hearing. Spiking Neural Networks (SNNs), as brain-inspired c…
cs.CV2025
EventZoom: A Progressive Approach to Event-Based Data Augmentation for Enhanced Neuromorphic Vision
Yiting Dong, Xiang He, Guobin Shen +3
Dynamic Vision Sensors (DVS) capture event data with high temporal resolution and low power consumption, presenting a more efficient solution for visual processing in dynamic and r…