1 citations · 1 across the 7 of their papers we have counts for
8 papers
Evidence for Dynamical Filtering: High Binary Fraction, Hard-binary Excess, and Unresolved Triples in the Surviving Core of NGC 6791
Huanbin Chi, Zhi Li, Feng Wang +4
We present a deep photometric analysis of the main-sequence (MS) population in the old, metal-rich open cluster (OC) NGC 6791 using Gaia Data Release 3 data. After correcting for d…
Search, Inspect, Fetch: Exploiting Structure-Aware Boolean Retrieval for Deep-Research Agents
Shuai Wang, Haodong Chen, Yu Yin +3
Existing deep-research agents use a Search--Visit workflow that retrieves whole webpages without considering the structure they expose through titles, headings, sections, and metad…
Whole-Pool Setwise Reranking with Long-Context Language Models
Hang Li, Chuting Yu, Teerapong Leelanupab +2
Previous LLM-based passage re-rankers are often expensive and slow because the input context constraints require the LLM to make many dependent model calls. We study how recent lon…
On the impact of retrieved content representations in RAG Pipelines
Jonathan J Ross, Bevan Koopman, Anton van der Vegt +1
Retrieval-Augmented Generation (RAG) supplements a language model's input with retrieved documents, yet most RAG pipelines inherit retrieval components designed for human readers.…
DiffRetriever: Parallel Representative Tokens for Retrieval with Diffusion Language Models
Shuai Wang, Yu Yin, Shengyao Zhuang +2
This paper shows how diffusion language models (DLMs) can be used as effective and efficient retrievers. Existing DLM-based retrievers (e.g., DiffEmbed) follow BERT-style encoding,…
When LLM Judges Inflate Scores: Exploring Overrating in Relevance Assessment
Chuting Yu, Hang Li, Guido Zuccon +2
Human relevance assessment is time-consuming and cognitively intensive, limiting the scalability of Information Retrieval evaluation. This has led to growing interest in using larg…