benchmark 1chain-of-thought prompting 1financial literacy 1numerical reasoning 1program-of-thought prompting 1
From the 1 of 20 linked papers with an AI index.
Showing cs.LGShow all
2 papers · 1 filter
cs.LG2026
Evolutionary Task Discovery: Advancing Reasoning Frontiers via Skill Composition and Complexity Scaling
Liqin Ye, Yanbin Yin, Michael Galarnyk +3
The reasoning frontier of Large Language Models (LLMs) has advanced significantly through modern post-training paradigms (e.g., Reinforcement Learning from Verifiable Rewards (RLVR…
cs.LG2025
Financial Instruction Following Evaluation (FIFE)
Glenn Matlin, Siddharth, Anirudh JM +3
Language Models (LMs) struggle with complex, interdependent instructions, particularly in high-stakes domains like finance where precision is critical. We introduce FIFE, a novel,…