A Survey of Multi-View Representation Learning
arXiv:1610.01206 · doi:10.1109/TKDE.2018.2872063
Abstract
Recently, multi-view representation learning has become a rapidly growing direction in machine learning and data mining areas. This paper introduces two categories for multi-view representation learning: multi-view representation alignment and multi-view representation fusion. Consequently, we first review the representative methods and theories of multi-view representation learning based on the perspective of alignment, such as correlation-based alignment. Representative examples are canonical correlation analysis (CCA) and its several extensions. Then from the perspective of representation fusion we investigate the advancement of multi-view representation learning that ranges from generative methods including multi-modal topic learning, multi-view sparse coding, and multi-view latent space Markov networks, to neural network-based methods including multi-modal autoencoders, multi-view convolutional neural networks, and multi-modal recurrent neural networks. Further, we also investigate several important applications of multi-view representation learning. Overall, this survey aims to provide an insightful overview of theoretical foundation and state-of-the-art developments in the field of multi-view representation learning and to help researchers find the most appropriate tools for particular applications.
Accepted by IEEE Transactions on Knowledge and Data Engineering
References in corpus (1)
Cited by in corpus (21)
- Memory-Guided Multi-View Multi-Domain Fake News Detection
- Imbalanced Big Data Oversampling: Taxonomy, Algorithms, Software, Guidelines and Future Directions
- A survey of multimodal deep generative models
- A Concise yet Effective model for Non-Aligned Incomplete Multi-view and Missing Multi-label Learning
- SleepPoseNet: Multi-View Learning for Sleep Postural Transition Recognition Using UWB
- A Clustering-guided Contrastive Fusion for Multi-view Representation Learning
- Semantic Invariant Multi-view Clustering with Fully Incomplete Information
- MP-SeizNet: A Multi-Path CNN Bi-LSTM Network for Seizure-Type Classification Using EEG
- Robust Multi-agent Communication via Multi-view Message Certification
- Common Practices and Taxonomy in Deep Multi-view Fusion for Remote Sensing Applications
- Deep Models for Multi-View 3D Object Recognition: A Review
- Muti-view Mouse Social Behaviour Recognition with Deep Graphical Model
- Disentangling Multi-view Representations Beyond Inductive Bias
- MORI-RAN: Multi-view Robust Representation Learning via Hybrid Contrastive Fusion
- A Temporal-Spectral Fusion Transformer with Subject-Specific Adapter for Enhancing RSVP-BCI Decoding
- Recent Advances and Challenges in Deep Audio-Visual Correlation Learning
- High-dimentional Multipartite Entanglement Structure Detection with Low Cost
- Deep Code Search with Naming-Agnostic Contrastive Multi-View Learning
- Generalized Multi-view Shared Subspace Learning using View Bootstrapping
- Multimodal Representation Learning using Deep Multiset Canonical Correlation
- Document Classification Pattern Recognition via Information Fusion: A Systematic Review of Multimodal and Multiview Representation Approaches