Machine Knowledge: Creation and Curation of Comprehensive Knowledge Bases
arXiv:2009.11564
Abstract
Equipping machines with comprehensive knowledge of the world's entities and their relationships has been a long-standing goal of AI. Over the last decade, large-scale knowledge bases, also known as knowledge graphs, have been automatically constructed from web contents and text sources, and have become a key asset for search engines. This machine knowledge can be harnessed to semantically interpret textual phrases in news, social media and web tables, and contributes to question answering, natural language processing and data analytics. This article surveys fundamental concepts and practical methods for creating and curating large knowledge bases. It covers models and methods for discovering and canonicalizing entities and their semantic types and organizing them into clean taxonomies. On top of this, the article discusses the automatic extraction of entity-centric properties. To support the long-term life-cycle and the quality assurance of machine knowledge, the article presents methods for constructing open schemas and for knowledge curation. Case studies on academic projects and industrial knowledge graphs complement the survey of concepts and methods.
Submitted to Foundations and Trends in Databases
References in corpus (20)
- Deep Learning in Neural Networks: An Overview
- A Survey on Knowledge Graphs: Representation, Acquisition and Applications
- Pre-trained Models for Natural Language Processing: A Survey
- Analysis of Named Entity Recognition and Linking for Tweets
- Improving the Accuracy and Efficiency of MAP Inference for Markov Logic
- Neural Entity Linking: A Survey of Models Based on Deep Learning
- Cross-Sentence N-ary Relation Extraction with Graph LSTMs
- Predicting Completeness in Knowledge Bases
- HoloClean: Holistic Data Repairs with Probabilistic Inference
- More Data, More Relations, More Context and More Openness: A Review and Outlook for Relation Extraction
- Knowledge Graphs on the Web -- an Overview
- Tuffy: Scaling up Statistical Inference in Markov Logic Networks using an RDBMS
- Generate FAIR Literature Surveys with Scholarly Knowledge Graphs
- Web Table Extraction, Retrieval and Augmentation: A Survey
- SciLens: Evaluating the Quality of Scientific News Articles Using Social Media and Scientific Literature Indicators
- Incremental Knowledge Base Construction Using DeepDive
- Requirements Analysis for an Open Research Knowledge Graph
- TableQnA: Answering List Intent Queries With Web Tables
- OpenKI: Integrating Open Information Extraction and Knowledge Bases with Relation Inference
- Technical Report: Optimizing Human Involvement for Entity Matching and Consolidation
Cited by in corpus (18)
- Construction of Knowledge Graphs: State and Challenges
- The Perils & Promises of Fact-checking with Large Language Models
- Recommender systems based on graph embedding techniques: A comprehensive review
- Extracting Cultural Commonsense Knowledge at Scale
- Saga: A Platform for Continuous Construction and Serving of Knowledge At Scale
- Language Models As or For Knowledge Bases
- An Ecosystem for Personal Knowledge Graphs: A Survey and Research Roadmap
- UnCommonSense: Informative Negative Knowledge about Everyday Concepts
- Relevant Entity Selection: Knowledge Graph Bootstrapping via Zero-Shot Analogical Pruning
- A Library Perspective on Nearly-Unsupervised Information Extraction Workflows in Digital Libraries
- A Conceptual Model for Attributions in Event-Centric Knowledge Graphs
- Native Execution of GraphQL Queries over RDF Graphs Using Multi-way Joins
- Relational World Knowledge Representation in Contextual Language Models: A Review
- Combining Knowledge Graphs and NLP to Analyze Instant Messaging Data in Criminal Investigations
- A Library Perspective on Supervised Text Processing in Digital Libraries: An Investigation in the Biomedical Domain
- KGPrune: a Web Application to Extract Subgraphs of Interest from Wikidata with Analogical Pruning
- Modeling Social Readers: Novel Tools for Addressing Reception from Online Book Reviews
- Commonsense Knowledge Base Construction in the Age of Big Data