2 papers
cs.AI2026
Mr.LHDR: A Benchmark for Multimodal Real-World Long-Horizon Deep Research Agents
Minghao Guo, Meng Cao, Sui Zhao +10
Deep research agents are increasingly capable of web search, tool use, multimodal evidence analysis, and information synthesis. However, existing benchmarks mainly evaluate medium-…
cs.IR2026
HeadRank: Decoding-Free Passage Reranking via Preference-Aligned Attention Heads
Juyuan Wang, Chenxing Wang, Yuchen Fang +8
Decoding-free reranking methods that read relevance signals directly from LLM attention weights offer significant latency advantages over autoregressive approaches, yet suffer from…