Reward Function Design for Crowd Simulation via Reinforcement Learning
arXiv:2309.12841 · doi:10.1145/3623264.3624452
Abstract
Crowd simulation is important for video-games design, since it enables to populate virtual worlds with autonomous avatars that navigate in a human-like manner. Reinforcement learning has shown great potential in simulating virtual crowds, but the design of the reward function is critical to achieving effective and efficient results. In this work, we explore the design of reward functions for reinforcement learning-based crowd simulation. We provide theoretical insights on the validity of certain reward functions according to their analytical properties, and evaluate them empirically using a range of scenarios, using the energy efficiency as the metric. Our experiments show that directly minimizing the energy usage is a viable strategy as long as it is paired with an appropriately scaled guiding potential, and enable us to study the impact of the different reward components on the behavior of the simulated crowd. Our findings can inform the development of new crowd simulation techniques, and contribute to the wider study of human-like navigation.
References in corpus (7)
- Multi-agent Reinforcement Learning in Sequential Social Dilemmas
- Implementation Matters in Deep Policy Gradients: A Case Study on PPO and TRPO
- What Matters In On-Policy Reinforcement Learning? A Large-Scale Empirical Study
- Hyperbolic Discounting and Learning over Multiple Horizons
- Discounted Reinforcement Learning Is Not an Optimization Problem
- A Survey on Reinforcement Learning Methods in Character Animation
- UGAE: A Novel Approach to Non-exponential Discounting