6 papers
Spiking Pyramid Wavelet Transformation for High-efficient and Low-energy Image Restoration
Chen Zhao, Xiantao Hu, Song Wu +5
Spiking neural networks (SNNs) have garnered significant interest in computer vision due to their potential for efficiency and biological inspiration. While spiking CNN-based metho…
From Contrast to Consistency: Rethinking Event-based Continuous-Time Optical Flow Estimation
Rui Hu, Song Wu, Wen Yang +1
Estimating continuous optical flow is a fundamental yet challenging problem in dynamic visual perception. Event-based cameras, with microsecond latency and high dynamic range, capt…
SimInsert: Seamless Video Object Insertion via Regional Sparse Attention Fusion
Xinyu Chen, Yuyi Qian, Jiang Lin +9
Video object insertion requires ensuring spatio-temporal coherence and interactive realism, extending far beyond simple content placement. However, current approaches are often hin…
TripVVT: A Large-Scale Triplet Dataset and a Coarse-Mask Baseline for In-the-Wild Video Virtual Try-On
Dingbao Shao, Song Wu, Shenyi Wang +9
Due to the scarcity of large-scale in-the-wild triplet data and the improper use of masks, the performance of video virtual try-on models remains limited. In this paper, we first i…
Tiny-YOLOSAM: Fast Hybrid Image Segmentation
Kenneth Xu, Songhan Wu
The Segment Anything Model (SAM) enables promptable, high-quality segmentation but is often too computationally expensive for latency-critical settings. TinySAM is a lightweight, d…
Traffic Scene Small Target Detection Method Based on YOLOv8n-SPTS Model for Autonomous Driving
Songhan Wu
This paper focuses on the key issue in autonomous driving: small target recognition in dynamic perception. Existing algorithms suffer from poor detection performance due to missing…