End-to-End Neural Ad-hoc Ranking with Kernel Pooling
arXiv:1706.06613 · doi:10.1145/3077136.3080809
Abstract
This paper proposes K-NRM, a kernel based neural model for document ranking. Given a query and a set of documents, K-NRM uses a translation matrix that models word-level similarities via word embeddings, a new kernel-pooling technique that uses kernels to extract multi-level soft match features, and a learning-to-rank layer that combines those features into the final ranking score. The whole model is trained end-to-end. The ranking layer learns desired feature patterns from the pairwise ranking loss. The kernels transfer the feature patterns into soft-match targets at each similarity level and enforce them on the translation matrix. The word embeddings are tuned accordingly so that they can produce the desired soft matches. Experiments on a commercial search engine's query log demonstrate the improvements of K-NRM over prior feature-based and neural-based states-of-the-art, and explain the source of K-NRM's advantage: Its kernel-guided embedding encodes a similarity metric tailored for matching query words to document words, and provides effective multi-level soft matches.
References in corpus (2)
Cited by in corpus (77)
- Deeper Text Understanding for IR with Contextual Neural Language Modeling
- Multi-Stage Document Ranking with BERT
- ERNIE 3.0: Large-scale Knowledge Enhanced Pre-training for Language Understanding and Generation
- Understanding the Behaviors of BERT in Ranking
- Simple Applications of BERT for Ad Hoc Document Retrieval
- Context-Aware Sentence/Passage Term Importance Estimation For First Stage Retrieval
- Word-Entity Duet Representations for Document Ranking
- Representation Learning for Natural Language Processing
- Pseudo-Relevance Feedback for Multiple Representation Dense Retrieval
- MatchZoo: A Learning, Practicing, and Developing System for Neural Text Matching
- The Expando-Mono-Duo Design Pattern for Text Ranking with Pretrained Sequence-to-Sequence Models
- A Study of Neural Matching Models for Cross-lingual IR
- StruBERT: Structure-aware BERT for Table Search and Matching
- An Updated Duet Model for Passage Re-ranking
- Learning a Product Relevance Model from Click-Through Data in E-Commerce
- Contrastive Learning of User Behavior Sequence for Context-Aware Document Ranking
- USER: A Unified Information Search and Recommendation Model based on Integrated Behavior Sequence
- Learning Implicit User Profiles for Personalized Retrieval-Based Chatbot
- Cross-lingual Information Retrieval with BERT
- Selective Weak Supervision for Neural Information Retrieval
- Incorporating Query Term Independence Assumption for Efficient Retrieval and Ranking using Deep Neural Networks
- Enhancing User Behavior Sequence Modeling by Generative Tasks for Session Search
- Passage Ranking with Weak Supervision
- Graph-based Hierarchical Relevance Matching Signals for Ad-hoc Retrieval
- Leveraging Semantic and Lexical Matching to Improve the Recall of Document Retrieval Systems: A Hybrid Approach
- P^3 Ranker: Mitigating the Gaps between Pre-training and Ranking Fine-tuning with Prompt-based Learning and Pre-finetuning
- A Deep Look into Neural Ranking Models for Information Retrieval
- TU Wien @ TREC Deep Learning '19 -- Simple Contextualization for Re-ranking
- Let's measure run time! Extending the IR replicability infrastructure to include performance aspects
- Learning Contextualized Document Representations for Healthcare Answer Retrieval
- GRAPHENE: A Precise Biomedical Literature Retrieval Engine with Graph Augmented Deep Learning and External Knowledge Empowerment
- COIL: Revisit Exact Lexical Match in Information Retrieval with Contextualized Inverted List
- Intra-Document Cascading: Learning to Select Passages for Neural Document Ranking
- Legal Element-oriented Modeling with Multi-view Contrastive Learning for Legal Case Retrieval
- From Easy to Hard: A Dual Curriculum Learning Framework for Context-Aware Document Ranking
- Improving Low-Resource Cross-lingual Document Retrieval by Reranking with Deep Bilingual Representations
- Co-PACRR: A Context-Aware Neural IR Model for Ad-hoc Retrieval
- One word at a time: adversarial attacks on retrieval models
- Longformer for MS MARCO Document Re-ranking Task
- A Modern Perspective on Query Likelihood with Deep Generative Retrieval Models
- LATTE: Latent Type Modeling for Biomedical Entity Linking
- SPARTA: Efficient Open-Domain Question Answering via Sparse Transformer Matching Retrieval
- A Comparison of Supervised Learning to Match Methods for Product Search
- Learning Colour Representations of Search Queries
- Dialogue History Matters! Personalized Response Selectionin Multi-turn Retrieval-based Chatbots
- Leveraging Advantages of Interactive and Non-Interactive Models for Vector-Based Cross-Lingual Information Retrieval
- Semantic Product Search for Matching Structured Product Catalogs in E-Commerce
- Heterogeneous Network Embedding for Deep Semantic Relevance Match in E-commerce Search
- Toward the Understanding of Deep Text Matching Models for Information Retrieval
- Developing Multi-Task Recommendations with Long-Term Rewards via Policy Distilled Reinforcement Learning
- Patient Cohort Retrieval using Transformer Language Models
- Beyond Lexical: A Semantic Retrieval Framework for Textual SearchEngine
- Learning to Select Historical News Articles for Interaction based Neural News Recommendation
- Neural Ranking Models with Multiple Document Fields
- GazBy: Gaze-Based BERT Model to Incorporate Human Attention in Neural Information Retrieval
- Pre-trained Language Model based Ranking in Baidu Search
- A multi-perspective combined recall and rank framework for Chinese procedure terminology normalization
- DeText: A Deep Text Ranking Framework with BERT
- Embedding-based Product Retrieval in Taobao Search
- Ember: No-Code Context Enrichment via Similarity-Based Keyless Joins
- Investigating Retrieval Method Selection with Axiomatic Features
- Improving Query Representations for Dense Retrieval with Pseudo Relevance Feedback
- Pre-training for Ad-hoc Retrieval: Hyperlink is Also You Need
- Mitigating the Position Bias of Transformer Models in Passage Re-Ranking
- LRG at TREC 2020: Document Ranking with XLNet-Based Models
- Exploiting Sentence-Level Representations for Passage Ranking
- BERT Embeddings Can Track Context in Conversational Search
- Place Deduplication with Embeddings
- Enriching Article Recommendation with Phrase Awareness
- You Get What You Chat: Using Conversations to Personalize Search-based Recommendations
- Cross-Batch Negative Sampling for Training Two-Tower Recommenders
- Deep Natural Language Processing for LinkedIn Search
- Target-Guided Open-Domain Conversation
- TPRM: A Topic-based Personalized Ranking Model for Web Search
- The Cross-Lingual Arabic Information REtrieval (CLAIRE) System
- ExpertRank: A Multi-level Coarse-grained Expert-based Listwise Ranking Loss
- More Robust Dense Retrieval with Contrastive Dual Learning