2 papers
cs.AI2026
Completion vs Optimality: Policy Gradient in Long-Horizon Cumulative-Damage Problems
Wolfgang Maass, Sabine Janzen
Long-horizon decision problems with cumulative damage couple locally attractive actions to globally adverse outcomes. We identify two orthogonal failure modes for policy-gradient m…
cs.LG2026
Evolving Afferent Architectures: Biologically-inspired Models for Damage-Avoidance Learning
Wolfgang Maass, Sabine Janzen, Prajvi Saxena +1
We introduce Afferent Learning, a framework that produces Computational Afferent Traces (CATs) as adaptive, internal risk signals for damage-avoidance learning. Inspired by biologi…