3 papers
cs.CV2026
LightAVSeg: Lightweight Audio-Visual Segmentation
Qing Zhong, Guodong Ding, Lingqiao Liu +3
Audio-Visual Segmentation (AVS) targets pixel level localization of sounding emitting objects in videos. However, existing models rely on dense cross-modal attention with quadratic…
cs.CV2025
A Temporal Modeling Framework for Video Pre-Training on Video Instance Segmentation
Qing Zhong, Peng-Tao Jiang, Wen Wang +3
Contemporary Video Instance Segmentation (VIS) methods typically adhere to a pre-train then fine-tune regime, where a segmentation model trained on images is fine-tuned on videos.…
cs.CV2024
OnlineTAS: An Online Baseline for Temporal Action Segmentation
Qing Zhong, Guodong Ding, Angela Yao
Temporal context plays a significant role in temporal action segmentation. In an offline setting, the context is typically captured by the segmentation network after observing the…