benchmark 1chain-of-thought prompting 1financial literacy 1numerical reasoning 1program-of-thought prompting 1
From the 1 of 19 linked papers with an AI index.
Showing cs.AIShow all
2 papers · 1 filter
cs.AI2026
Precise Attribute Intensity Control in Large Language Models via Targeted Representation Editing
Rongzhi Zhang, Liqin Ye, Yuzhao Heng +5
Precise attribute intensity control--generating Large Language Model (LLM) outputs with specific, user-defined attribute intensities--is crucial for AI systems adaptable to diverse…
cs.AI2026
FinForge: Semi-Synthetic Financial Benchmark Generation
Glenn Matlin, Akhil Theerthala, Anant Gupta +4
Evaluating Language Models (LMs) in specialized, high-stakes domains such as finance remains a significant challenge due to the scarcity of open, high-quality, and domain-specific…