Content-based Video Indexing and Retrieval Using Corr-LDA
arXiv:1602.08581
Abstract
Existing video indexing and retrieval methods on popular web-based multimedia sharing websites are based on user-provided sparse tagging. This paper proposes a very specific way of searching for video clips, based on the content of the video. We present our work on Content-based Video Indexing and Retrieval using the Correspondence-Latent Dirichlet Allocation (corr-LDA) probabilistic framework. This is a model that provides for auto-annotation of videos in a database with textual descriptors, and brings the added benefit of utilizing the semantic relations between the content of the video and text. We use the concept-level matching provided by corr-LDA to build correspondences between text and multimedia, with the objective of retrieving content with increased accuracy. In our experiments, we employ only the audio components of the individual recordings and compare our results with an SVM-based approach.
8 Pages, Updated References, Added Figures
References in corpus (4)
Cited by in corpus (9)
- An Unsupervised Domain-Independent Framework for Automated Detection of Persuasion Tactics in Text
- Event Outcome Prediction using Sentiment Analysis and Crowd Wisdom in Microblog Feeds
- A Machine Learning Framework for Authorship Identification From Texts
- A Heterogeneous Graphical Model to Understand User-Level Sentiments in Social Media
- Simultaneous Identification of Tweet Purpose and Position
- Modeling Product Search Relevance in e-Commerce
- A Correspondence Analysis Framework for Author-Conference Recommendations
- An End-to-End ML System for Personalized Conversational Voice Models in Walmart E-Commerce
- Transition-Based Dependency Parsing using Perceptron Learner