collaborators

6 papers

cs.CV2026

Spiking Pyramid Wavelet Transformation for High-efficient and Low-energy Image Restoration

Chen Zhao, Xiantao Hu, Song Wu +5

Spiking neural networks (SNNs) have garnered significant interest in computer vision due to their potential for efficiency and biological inspiration. While spiking CNN-based metho…

cs.CV2026

From Contrast to Consistency: Rethinking Event-based Continuous-Time Optical Flow Estimation

Rui Hu, Song Wu, Wen Yang +1

Estimating continuous optical flow is a fundamental yet challenging problem in dynamic visual perception. Event-based cameras, with microsecond latency and high dynamic range, capt…

cs.CV2026

SimInsert: Seamless Video Object Insertion via Regional Sparse Attention Fusion

Xinyu Chen, Yuyi Qian, Jiang Lin +9

Video object insertion requires ensuring spatio-temporal coherence and interactive realism, extending far beyond simple content placement. However, current approaches are often hin…

cs.CV2026

TripVVT: A Large-Scale Triplet Dataset and a Coarse-Mask Baseline for In-the-Wild Video Virtual Try-On

Dingbao Shao, Song Wu, Shenyi Wang +9

Due to the scarcity of large-scale in-the-wild triplet data and the improper use of masks, the performance of video virtual try-on models remains limited. In this paper, we first i…

cs.CV2025

Tiny-YOLOSAM: Fast Hybrid Image Segmentation

Kenneth Xu, Songhan Wu

The Segment Anything Model (SAM) enables promptable, high-quality segmentation but is often too computationally expensive for latency-critical settings. TinySAM is a lightweight, d…

cs.CV2025

Traffic Scene Small Target Detection Method Based on YOLOv8n-SPTS Model for Autonomous Driving

Songhan Wu

This paper focuses on the key issue in autonomous driving: small target recognition in dynamic perception. Existing algorithms suffer from poor detection performance due to missing…