3 papers
cs.CV2026
TripVVT: A Large-Scale Triplet Dataset and a Coarse-Mask Baseline for In-the-Wild Video Virtual Try-On
Dingbao Shao, Song Wu, Shenyi Wang +9
Due to the scarcity of large-scale in-the-wild triplet data and the improper use of masks, the performance of video virtual try-on models remains limited. In this paper, we first i…
cs.CV2025
Tiny-YOLOSAM: Fast Hybrid Image Segmentation
Kenneth Xu, Songhan Wu
The Segment Anything Model (SAM) enables promptable, high-quality segmentation but is often too computationally expensive for latency-critical settings. TinySAM is a lightweight, d…
cs.CV2025
Traffic Scene Small Target Detection Method Based on YOLOv8n-SPTS Model for Autonomous Driving
Songhan Wu
This paper focuses on the key issue in autonomous driving: small target recognition in dynamic perception. Existing algorithms suffer from poor detection performance due to missing…