39 citations · 40 across the 2 of their papers we have counts for
2 papers
cs.CV2023★ 1 cited
Joint Representation Learning for Text and 3D Point Cloud
Rui Huang, Xuran Pan, Henry Zheng +4
Recent advancements in vision-language pre-training (e.g. CLIP) have shown that vision models can benefit from language supervision. While many models using language modality have…
cs.CV2022★ 39 cited
Boosting Night-time Scene Parsing with Learnable Frequency
Zhifeng Xie, Sen Wang, Ke Xu +4
Night-Time Scene Parsing (NTSP) is essential to many vision applications, especially for autonomous driving. Most of the existing methods are proposed for day-time scene parsing. T…