2 papers
cs.AI2026
Deep FinResearch Bench: Evaluating AI's Ability to Conduct Professional Financial Investment Research
Mirazul Haque, Antony Papadimitriou, Samuel Mensah +6
We introduce Deep FinResearch Bench, a practical and comprehensive evaluation framework for deep research (DR) agents in financial investment research. The benchmark assesses three…
cs.CL2026
Entropy-Gated Branching for Efficient Test-Time Reasoning
Xianzhi Li, Ethan Callanan, Abdellah Ghassel +1
Test-time compute methods can significantly improve the reasoning capabilities and problem-solving accuracy of large language models (LLMs). However, these approaches require subst…