Searching Personal Collections
arXiv:2412.12330 · doi:10.1145/3674127.3674142
Abstract
This article describes the history of information retrieval on personal document collections.
References in corpus (123)
- Power laws, Pareto distributions and Zipf's law
- Google's Neural Machine Translation System: Bridging the Gap between Human and Machine Translation
- LLaMA: Open and Efficient Foundation Language Models
- Exploring the Limits of Transfer Learning with a Unified Text-to-Text Transformer
- Llama 2: Open Foundation and Fine-Tuned Chat Models
- Longformer: The Long-Document Transformer
- Domain-Specific Language Model Pretraining for Biomedical Natural Language Processing
- Knowledge Graphs
- Evaluating Large Language Models Trained on Code
- Exploiting Similarities among Languages for Machine Translation
- Recurrent Neural Networks with Top-k Gains for Session-based Recommendations
- A Deep Relevance Matching Model for Ad-hoc Retrieval
- Prediction-Based Decisions and Fairness: A Catalogue of Choices, Assumptions, and Definitions
- Fairness of Exposure in Rankings
- End-to-End Neural Ad-hoc Ranking with Kernel Pooling
- Beyond the Imitation Game: Quantifying and extrapolating the capabilities of language models
- FA*IR: A Fair Top-k Ranking Algorithm
- Think you have Solved Question Answering? Try ARC, the AI2 Reasoning Challenge
- The Microsoft 2017 Conversational Speech Recognition System
- Causal Intervention for Leveraging Popularity Bias in Recommendation
- A Survey on Conversational Recommender Systems
- A Survey on Accuracy-oriented Neural Recommendation: From Collaborative Filtering to Information-rich Recommendation
- Deeper Text Understanding for IR with Contextual Neural Language Modeling
- How Algorithmic Confounding in Recommendation Systems Increases Homogeneity and Decreases Utility
- Equity of Attention: Amortizing Individual Fairness in Rankings
- Beyond Personalization: Research Directions in Multistakeholder Recommendation
- Massively Multilingual Word Embeddings
- Embarrassingly Shallow Autoencoders for Sparse Data
- A Survey Of Cross-lingual Word Embedding Models
- Measuring the Business Value of Recommender Systems
- Fairness in Recommender Systems: Research Landscape and Future Directions
- Document Expansion by Query Prediction
- Multi-Stage Document Ranking with BERT
- A Troubling Analysis of Reproducibility and Progress in Recommender Systems Research
- Unbiased Learning to Rank with Unbiased Propensity Estimation
- Research Commentary on Recommendations with Side Information: A Survey and Research Directions
- Evaluating Stochastic Rankings with Expected Exposure
- Can Generalist Foundation Models Outcompete Special-Purpose Tuning? Case Study in Medicine
- Perspectives on Large Language Models for Relevance Judgment
- How algorithmic popularity bias hinders or promotes quality
- Point process modeling for directed interaction networks
- Response Ranking with Deep Matching Networks and External Knowledge in Information-seeking Conversation Systems
- Conscientious Classification: A Data Scientist's Guide to Discrimination-Aware Classification
- Quizz: Targeted crowdsourcing with a billion (potential) users
- Empirical Analysis of Session-Based Recommendation Algorithms
- On Application of Learning to Rank for E-Commerce Search
- Learning Features of Music from Scratch
- Memento: Time Travel for the Web
- Visual Search at eBay
- Context-Aware Sentence/Passage Term Importance Estimation For First Stage Retrieval
- An Efficient Bandit Algorithm for Realtime Multivariate Optimization
- Modeling Diverse Relevance Patterns in Ad-hoc Retrieval
- Analyzing and Characterizing User Intent in Information-seeking Conversations
- Towards Conversational Diagnostic AI
- Fairness in Ranking: A Survey
- Patent Retrieval: A Literature Review
- SPICE: Self-supervised Pitch Estimation
- PatentBERT: Patent Classification with Fine-Tuning a pre-trained BERT Model
- A Study of MatchPyramid Models on Ad-hoc Retrieval
- A Graph-based Approach for Mitigating Multi-sided Exposure Bias in Recommender Systems
- In ChatGPT We Trust? Measuring and Characterizing the Reliability of ChatGPT
- Dynamical phase transitions, temporal orthogonality and the dynamics of observables in one dimensional ultra-cold quantum gases: from the continuum to the lattice
- Balancing Consumer and Business Value of Recommender Systems: A Simulation-based Analysis
- AutoKnow: Self-Driving Knowledge Collection for Products of Thousands of Types
- A Few Brief Notes on DeepImpact, COIL, and a Conceptual Framework for Information Retrieval Techniques
- Session-aware Recommendation: A Surprising Quest for the State-of-the-art
- Unifying Online and Counterfactual Learning to Rank
- ATBRG: Adaptive Target-Behavior Relational Graph Network for Effective Recommendation
- Retrieving Supporting Evidence for Generative Question Answering
- AGIEval: A Human-Centric Benchmark for Evaluating Foundation Models
- Improving Outfit Recommendation with Co-supervision of Fashion Generation
- SparTerm: Learning Term-based Sparse Representation for Fast Text Retrieval
- An Axiomatic Analysis of Diversity Evaluation Metrics: Introducing the Rank-Biased Utility Metric
- A Stakeholder-Centered View on Fairness in Music Recommender Systems
- The Information Retrieval Experiment Platform
- A Utility-Theoretic Approach to Privacy in Online Services
- Wisdom of the Crowd or Wisdom of a Few? An Analysis of Users' Content Generation
- Rewiring What-to-Watch-Next Recommendations to Reduce Radicalization Pathways
- Fast Passage Re-ranking with Contextualized Exact Term Matching and Efficient Passage Expansion
- Meta-evaluation of Conversational Search Evaluation Metrics
- Evaluating Generative Ad Hoc Information Retrieval
- How well do LLMs cite relevant medical references? An evaluation framework and analyses
- Attentive Memory Networks: Efficient Machine Reading for Conversational Search
- Query Embedding Pruning for Dense Retrieval
- A Study of Neural Matching Models for Cross-lingual IR
- Measuring Disparate Outcomes of Content Recommendation Algorithms with Distributional Inequality Metrics
- Deep Learning for Singing Processing: Achievements, Challenges and Impact on Singers and Listeners
- Explainable Information Retrieval: A Survey
- C3: Continued Pretraining with Contrastive Weak Supervision for Cross Language Ad-Hoc Retrieval
- Evaluation Measures of Individual Item Fairness for Recommender Systems: A Critical Study
- Studying the Transfer of Biases from Programmers to Programs
- ir_metadata: An Extensible Metadata Schema for IR Experiments
- Making Neural Networks Interpretable with Attribution: Application to Implicit Signals Prediction
- ViTOR: Learning to Rank Webpages Based on Visual Features
- Perceptions of Diversity in Electronic Music: the Impact of Listener, Artist, and Track Characteristics
- HAGRID: A Human-LLM Collaborative Dataset for Generative Information-Seeking with Attribution
- Constraint Translation Candidates: A Bridge between Neural Query Translation and Cross-lingual Information Retrieval
- Exploiting Neural Query Translation into Cross Lingual Information Retrieval
- Modeling Rabbit-Holes on YouTube
- Proceedings of the 17th Dutch-Belgian Information Retrieval Workshop
- Entertaining and Opinionated but Too Controlling: A Large-Scale User Study of an Open Domain Alexa Prize System
- Mixed Attention Transformer for Leveraging Word-Level Knowledge to Neural Cross-Lingual Information Retrieval
- A New Email Retrieval Ranking Approach
- A New Email Retrieval Ranking Approach
- Early MFCC And HPCP Fusion for Robust Cover Song Identification
- Principled Multi-Aspect Evaluation Measures of Rankings
- TripJudge: A Relevance Judgement Test Collection for TripClick Health Retrieval
- Report from Dagstuhl Seminar 23031: Frontiers of Information Access Experimentation for Research and Education
- EENMF: An End-to-End Neural Matching Framework for E-Commerce Sponsored Search
- Information search in a professional context - exploring a collection of professional search tasks
- Response to Moffat's Comment on "Towards Meaningful Statements in IR Evaluation: Mapping Evaluation Measures to Interval Scales"
- Usability Evaluation for Online Professional Search in the Dutch Archaeology Domain
- Helping results assessment by adding explainable elements to the deep relevance matching model
- Contrastive language and vision learning of general fashion concepts
- A Comparison of Methods for Evaluating Generative IR
- Unbiased Cascade Bandits: Mitigating Exposure Bias in Online Learning to Rank Recommendation
- Social network modeling and applications, a tutorial
- Result Diversification in Search and Recommendation: A Survey
- How Discriminative Are Your Qrels? How To Study the Statistical Significance of Document Adjudication Methods
- Cover Detection using Dominant Melody Embeddings
- Towards Verifiable Generation: A Benchmark for Knowledge-aware Language Model Attribution
- BiTimeBERT: Extending Pre-Trained Language Representations with Bi-Temporal Information
- Categorical, Ratio, and Professorial Data: The Case for Reciprocal Rank