Showing cs.LGShow all
2 papers · 1 filter
cs.LG2026
A Scalable Approach to Solving Simulation-Based Network Security Games
Michael Lanier, Yevgeniy Vorobeychik
We introduce MetaDOAR, a lightweight meta-controller that augments the Double Oracle / PSRO paradigm with a learned, partition-aware filtering layer and Q-value caching to enable s…
cs.LG2025
Learning Policy Committees for Effective Personalization in MDPs with Diverse Tasks
Luise Ge, Michael Lanier, Anindya Sarkar +3
Many dynamic decision problems, such as robotic control, involve a series of tasks, many of which are unknown at training time. Typical approaches for these problems, such as multi…