Learning Robust Video Synchronization without Annotations
arXiv:1610.05985
Abstract
Aligning video sequences is a fundamental yet still unsolved component for a broad range of applications in computer graphics and vision. Most classical image processing methods cannot be directly applied to related video problems due to the high amount of underlying data and their limit to small changes in appearance. We present a scalable and robust method for computing a non-linear temporal video alignment. The approach autonomously manages its training data for learning a meaningful representation in an iterative procedure each time increasing its own knowledge. It leverages on the nature of the videos themselves to remove the need for manually created labels. While previous alignment methods similarly consider weather conditions, season and illumination, our approach is able to align videos from data recorded months apart.
International Conference On Machine Learning And Applications (ICMLA 2017)
References in corpus (6)
- Very Deep Convolutional Networks for Large-Scale Image Recognition
- Theano: new features and speed improvements
- Learning to Compare Image Patches via Convolutional Neural Networks
- Stacked Attention Networks for Image Question Answering
- WarpNet: Weakly Supervised Matching for Single-view Reconstruction
- The Video Genome