2 papers
cs.LG2026
Learning to Control Coupled-Dynamics Environments with Joint Markov Decision Processes
Ege C. Kaya, Aliasghar Pourghani, Mahsa Ghasemi +2
Coupled-dynamics environments expose the one-step outcomes that would follow from several possible counterfactual actions under a common realization of exogenous randomness. The or…
cs.LG2026
Quotient-Categorical Representations for Bellman-Compatible Average-Reward Distributional Reinforcement Learning
Ege C. Kaya, Aliasghar Pourghani, Vijay Gupta +1
Average-reward reinforcement learning requires estimating the gain and the bias, which is defined only up to an additive constant. This makes direct distributional analogues ill-po…