1 citations · 1 across the 12 of their papers we have counts for
6 papers · 1 filter
Effective Reinforcement Learning for Agentic Search by Recycling Zero-Variance Queries During Training
João Coelho, João Magalhães, Bruno Martins +1
The use of GRPO-style algorithms has become the standard strategy for training LLM search agents under outcome-only rewards. With these algorithms, a query contributes to parameter…
Efficient Dataset Selection for Continual Adaptation of Generative Recommenders
Cathy Jiao, Juan Elenter, Praveen Ravichandran +7
Recommendation systems must continuously adapt to evolving user behavior, yet the volume of data generated in large-scale streaming environments makes frequent full retraining impr…
Agentic Search in the Wild: Intents and Trajectory Dynamics from 14M+ Real Search Requests
Jingjie Ning, João Coelho, Yibo Kong +5
LLM-powered search agents are increasingly being used for multi-step information seeking tasks, yet the IR community lacks empirical understanding of how agentic search sessions un…
Aligning Web Query Generation with Ranking Objectives via Direct Preference Optimization
João Coelho, Bruno Martins, João Magalhães +1
Neural retrieval models excel in Web search, but their training requires substantial amounts of labeled query-document pairs, which are costly to obtain. With the widespread availa…
DeepResearchGym: A Free, Transparent, and Reproducible Evaluation Sandbox for Deep Research
João Coelho, Jingjie Ning, Jingyuan He +8
Deep research systems represent an emerging class of agentic information retrieval methods that generate comprehensive and well-supported reports to complex queries. However, most…
Dwell in the Beginning: How Language Models Embed Long Documents for Dense Retrieval
João Coelho, Bruno Martins, João Magalhães +2
This study investigates the existence of positional biases in Transformer-based models for text representation learning, particularly in the context of web document retrieval. We b…