Showing cs.LGShow all
2 papers · 1 filter
cs.LG2024
Accelerating Proximal Policy Optimization Learning Using Task Prediction for Solving Environments with Delayed Rewards
Ahmad Ahmad, Mehdi Kermanshah, Kevin Leahy +6
In this paper, we tackle the challenging problem of delayed rewards in reinforcement learning (RL). While Proximal Policy Optimization (PPO) has emerged as a leading Policy Gradien…
cs.LG2024
Learning Optimal Signal Temporal Logic Decision Trees for Classification: A Max-Flow MILP Formulation
Kaier Liang, Gustavo A. Cardona, Disha Kamale +1
This paper presents a novel framework for inferring timed temporal logic properties from data. The dataset comprises pairs of finite-time system traces and corresponding labels, de…