4 papers
Processing Network Controls via Deep Reinforcement Learning
Mark Gluzman
Novel advanced policy gradient (APG) algorithms, such as proximal policy optimization (PPO), trust region policy optimization, and their variations, have become the dominant reinfo…
Refined Policy Improvement Bounds for MDPs
J. G. Dai, Mark Gluzman
The policy improvement bound on the difference of the discounted returns plays a crucial role in the theoretical justification of the trust-region policy optimization (TRPO) algori…
Optimizing adaptive cancer therapy: dynamic programming and evolutionary game theory
Mark Gluzman, Jacob G. Scott, Alexander Vladimirsky
Recent clinical trials have shown that the adaptive drug therapy can be more efficient than a standard MTD-based policy in treatment of cancer patients. The adaptive therapy paradi…
Dynamics of Two Coupled van der Pol Oscillators with Delay Coupling Revisited
Mark Gluzman, Richard Rand
The problem of two van der Pol oscillators coupled by velocity delay terms was studied by Wirkus and Rand in 2002. The small-epsilon analysis resulted in a slow flow which containe…