agentic search 1context management 1information retrieval 1long-horizon tasks 1web browsing agents 1
From the 1 of 3 linked papers with an AI index.
3 papers
cs.AI2026
FrontierFinance: A Challenging Benchmark for Measuring Frontier Intelligence of Finance Agents
Yuhao Zhang, O. Ozan Koyluoglu, Thejas Venkatesh +4
AI agents are increasingly deployed for professional investment research, yet no benchmark captures the complexity of the full investor workflow. Existing benchmarks mainly target…
cs.CL2026
Lost in the Maze: Overcoming Context Limitations in Long-Horizon Agentic Search
Howard Yen, Yoonsang Lee, Ashwin Paranjape +5
The paper introduces SLIM, a lightweight framework that separates search and browsing tools and periodically summarizes information to overcome context limits in long-horizon web‑a…
cs.CL2024
RAG-QA Arena: Evaluating Domain Robustness for Long-form Retrieval Augmented Question Answering
Rujun Han, Yuhao Zhang, Peng Qi +6
Question answering based on retrieval augmented generation (RAG-QA) is an important research topic in NLP and has a wide range of real-world applications. However, most existing da…