76 citations · 153 across the 21 of their papers we have counts for
22 papers · 1 filter
Open-vocabulary Semantic Segmentation with Frozen Vision-Language Models
Chaofan Ma, Yuhuan Yang, Yanfeng Wang +2
When trained at a sufficient scale, self-supervised learning has exhibited a notable ability to solve a wide range of visual or language understanding tasks. In this paper, we inve…
Number-Adaptive Prototype Learning for 3D Point Cloud Semantic Segmentation
Yangheng Zhao, Jun Wang, Xiaolong Li +4
3D point cloud semantic segmentation is one of the fundamental tasks for 3D scene understanding and has been widely used in the metaverse applications. Many recent 3D semantic segm…
A Simple Plugin for Transforming Images to Arbitrary Scales
Qinye Zhou, Ziyi Li, Weidi Xie +3
Existing models on super-resolution often specialized for one scale, fundamentally limiting their use in practical scenarios. In this paper, we aim to develop a general plugin that…
Self-Supervised Masking for Unsupervised Anomaly Detection and Localization
Chaoqin Huang, Qinwei Xu, Yanfeng Wang +2
Recently, anomaly detection and localization in multimedia data have received significant attention among the machine learning community. In real-world applications such as medical…
Multiscale Spatio-Temporal Graph Neural Networks for 3D Skeleton-Based Motion Prediction
Maosen Li, Siheng Chen, Yangheng Zhao +3
We propose a multiscale spatio-temporal graph neural network (MST-GNN) to predict the future 3D skeleton-based human poses in an action-category-agnostic manner. The core of MST-GN…
CaT: Weakly Supervised Object Detection with Category Transfer
Tianyue Cao, Lianyu Du, Xiaoyun Zhang +3
A large gap exists between fully-supervised object detection and weakly-supervised object detection. To narrow this gap, some methods consider knowledge transfer from additional fu…