8 citations · 8 across the 2 of their papers we have counts for
2 papers
cs.LG2025
Wasserstein Policy Optimization
David Pfau, Ian Davies, Diana Borsa +3
We introduce Wasserstein Policy Optimization (WPO), an actor-critic algorithm for reinforcement learning in continuous action spaces. WPO can be derived as an approximation to Wass…
physics.plasm-ph2024★ 8 cited
TORAX: A Fast and Differentiable Tokamak Transport Simulator in JAX
Jonathan Citrin, Ian Goodfellow, Akhil Raju +11
We present TORAX, a new, open-source, differentiable tokamak core transport simulator implemented in Python using the JAX framework. TORAX solves the coupled equations for ion heat…