16 citations · 16 across the 3 of their papers we have counts for
3 papers
cs.CV2025
Frequency-Domain Decomposition and Recomposition for Robust Audio-Visual Segmentation
Yunzhe Shen, Kai Peng, Leiye Liu +5
Audio-visual segmentation (AVS) plays a critical role in multimodal machine learning by effectively integrating audio and visual cues to precisely segment objects or regions within…
cs.CV2025
SpikeVideoFormer: An Efficient Spike-Driven Video Transformer with Hamming Attention and Complexity
Shihao Zou, Qingfeng Li, Wei Ji +4
Spiking Neural Networks (SNNs) have shown competitive performance to Artificial Neural Networks (ANNs) in various vision tasks, while offering superior energy efficiency. However,…
cs.CV2022★ 16 cited
Promoting Saliency From Depth: Deep Unsupervised RGB-D Saliency Detection
Wei Ji, Jingjing Li, Qi Bi +3
Growing interests in RGB-D salient object detection (RGB-D SOD) have been witnessed in recent years, owing partly to the popularity of depth sensors and the rapid progress of deep…