3 papers
cs.CV2026
Bootstrapping Vision-Language Model for Hysteroscopic Surgical Scene Segmentation
Jun Huang, Meiyi Chen, Zijie Yue +6
Hysteroscopic surgical scene segmentation plays a pivotal role in understanding the hysteroscopic intraoperative environment as well as computer-assisted intervention. However, thi…
cs.CV2026
Weakly-Supervised Referring Video Object Segmentation through Text Supervision
Miaojing Shi, Jun Huang, Zijie Yue +1
Referring video object segmentation (RVOS) aims to segment the target instance in a video, referred by a text expression. Conventional approaches are mostly supervised learning, re…
cs.MM2024
TSC-PCAC: Voxel Transformer and Sparse Convolution Based Point Cloud Attribute Compression for 3D Broadcasting
Zixi Guo, Yun Zhang, Linwei Zhu +2
Point cloud has been the mainstream representation for advanced 3D applications, such as virtual reality and augmented reality. However, the massive data amounts of point clouds is…