Generalized Multi-view Shared Subspace Learning using View Bootstrapping
arXiv:2005.06038 · doi:10.1109/TSP.2021.3102751
Abstract
A key objective in multi-view learning is to model the information common to multiple parallel views of a class of objects/events to improve downstream learning tasks. In this context, two open research questions remain: How can we model hundreds of views per event? Can we learn robust multi-view embeddings without any knowledge of how these views are acquired? We present a neural method based on multi-view correlation to capture the information shared across a large number of views by subsampling them in a view-agnostic manner during training. To provide an upper bound on the number of views to subsample for a given embedding dimension, we analyze the error of the bootstrapped multi-view correlation objective using matrix concentration theory. Our experiments on spoken word recognition, 3D object classification and pose-invariant face recognition demonstrate the robustness of view bootstrapping to model a large number of views. Results underscore the applicability of our method for a view-agnostic learning setting.
References in corpus (10)
- Speech Commands: A Dataset for Limited-Vocabulary Speech Recognition
- A Survey on Multi-view Learning
- A kernel method for canonical correlation analysis
- Multi-View Spectral Clustering via Structured Low-Rank Matrix Factorization
- Generalizing Across Domains via Cross-Gradient Training
- Iterative Views Agreement: An Iterative Low-Rank based Structured Optimization Method to Multi-View Spectral Clustering
- DCASE 2018 Challenge - Task 5: Monitoring of domestic activities based on multi-channel acoustics
- Multiview Aggregation for Learning Category-Specific Shape Reconstruction
- Correlated Components Analysis - Extracting Reliable Dimensions in Multivariate Data
- Mapping individual differences in cortical architecture using multi-view representation learning