2 papers
cs.CV2025
TSMS-SAM2: Multi-scale Temporal Sampling Augmentation and Memory-Splitting Pruning for Promptable Video Object Segmentation and Tracking in Surgical Scenarios
Guoping Xu, Hua-Chieh Shao, You Zhang
Promptable video object segmentation and tracking (VOST) has seen significant advances with the emergence of foundation models like Segment Anything Model 2 (SAM2); however, their…
cs.CL2024
Personalized LoRA for Human-Centered Text Understanding
You Zhang, Jin Wang, Liang-Chih Yu +2
Effectively and efficiently adapting a pre-trained language model (PLM) for human-centered text understanding (HCTU) is challenging since user tokens are million-level in most pers…