2 papers
cs.CV2024
Enhancing Weakly Supervised Semantic Segmentation with Multi-modal Foundation Models: An End-to-End Approach
Elham Ravanbakhsh, Cheng Niu, Yongqing Liang +2
Semantic segmentation is a core computer vision problem, but the high costs of data annotation have hindered its wide application. Weakly-Supervised Semantic Segmentation (WSSS) of…
cs.CV2024
Deep video representation learning: a survey
Elham Ravanbakhsh, Yongqing Liang, J. Ramanujam +1
This paper provides a review on representation learning for videos. We classify recent spatiotemporal feature learning methods for sequential visual data and compare their pros and…