2 papers
cs.SD2025
Semantics-Aware Human Motion Generation from Audio Instructions
Zi-An Wang, Shihao Zou, Shiyao Yu +2
Recent advances in interactive technologies have highlighted the prominence of audio signals for semantic encoding. This paper explores a new task, where audio signals are used as…
cs.CV2025
SpikeVideoFormer: An Efficient Spike-Driven Video Transformer with Hamming Attention and Complexity
Shihao Zou, Qingfeng Li, Wei Ji +4
Spiking Neural Networks (SNNs) have shown competitive performance to Artificial Neural Networks (ANNs) in various vision tasks, while offering superior energy efficiency. However,…