Showing cs.AIShow all
2 papers · 1 filter
cs.AI2026
Evidence-State Rewards for Long-Context Reasoning
Ya Gao, Pekka Marttinen
Long-context reasoning requires models to locate, revise, and synthesize evidence distributed across lengthy inputs. Existing long-context RL methods usually reward final answers o…
cs.AI2026
Edit Knowledge, Not Just Facts via Multi-Step Reasoning over Background Stories
Ya Gao, Kalle Kujanpää, Pekka Marttinen +2
Enabling artificial intelligence systems, particularly large language models, to update knowledge and flexibly apply it during reasoning remains a central challenge. Existing knowl…