2 papers
cs.LG2026
LocusRL: Diagnosing LLM Reward and Policy Interventions in Competitive Games
Chengyu Luan, Bo Xin, Songyan Guo +4
Large language models can intervene in reinforcement learning through both reward design and action selection, yet aggregate performance offers an incomplete account of what these…
cs.AI2026
TimeNet: An Extensible Unified Data Infrastructure for Next-Generation Temporal Foundation Models
Martin Maritsch, Timo Stoffregen, Thomas Kaar +36
Temporal Foundation Models (TFMs) aim to generalize across domains, datasets, and tasks. Yet, their development remains constrained by fragmented, task-specific data formats, annot…