2 papers
cs.LG2026
Revisiting Action Factorization for Complex Action Spaces
Timothy Flavin, Sandip Sen
Many real-world control problems involve hybrid discrete-continuous action spaces. For example, steering and signaling in autonomous driving, and aiming and firing in robotics or v…
cs.MA2026
A High-Throughput Compute-Efficient POMDP Hide-And-Seek-Engine (HASE) for Multi-Agent Operations
Timothy Flavin, Sandip Sen
Reinforcement Learning (RL) algorithms exhibit high sample complexity, particularly when applied to Decentralized Partially Observable Markov Decision Processes (Dec-POMDPs). As a…