activity
20242026
collaborators

7 papers

cs.AI2026

Beyond Parallel Sampling: Diverse Query Initialization for Agentic Search

Sidhaarth Murali, João Coelho, Jingjie Ning +3

Test-time scaling for agentic search typically increases depth (i.e., more turns and tokens per trajectory) or breadth (i.e., more parallel rollouts). Here we focus on breadth scal…

cs.IR2026

Effective Reinforcement Learning for Agentic Search by Recycling Zero-Variance Queries During Training

João Coelho, João Magalhães, Bruno Martins +1

The use of GRPO-style algorithms has become the standard strategy for training LLM search agents under outcome-only rewards. With these algorithms, a query contributes to parameter…

cs.IR2026

Agentic Search in the Wild: Intents and Trajectory Dynamics from 14M+ Real Search Requests

Jingjie Ning, João Coelho, Yibo Kong +5

LLM-powered search agents are increasingly being used for multi-step information seeking tasks, yet the IR community lacks empirical understanding of how agentic search sessions un…

cs.IR2025

DeepResearchGym: A Free, Transparent, and Reproducible Evaluation Sandbox for Deep Research

João Coelho, Jingjie Ning, Jingyuan He +8

Deep research systems represent an emerging class of agentic information retrieval methods that generate comprehensive and well-supported reports to complex queries. However, most…

cs.CL2025

TeamCMU at Touché: Adversarial Co-Evolution for Advertisement Integration and Detection in Conversational Search

To Eun Kim, João Coelho, Gbemileke Onilude +1

As conversational search engines increasingly adopt generation-based paradigms powered by Large Language Models (LLMs) and Retrieval-Augmented Generation (RAG), the integration of…

cs.IR2025

Aligning Web Query Generation with Ranking Objectives via Direct Preference Optimization

João Coelho, Bruno Martins, João Magalhães +1

Neural retrieval models excel in Web search, but their training requires substantial amounts of labeled query-document pairs, which are costly to obtain. With the widespread availa…