4 papers
The PokeAgent Challenge: Competitive and Long-Context Learning at Scale
Seth Karten, Jake Grigsby, Tersoo Upaa +28
We present the PokeAgent Challenge, a large-scale benchmark for decision-making research built on Pokemon's multi-agent battle system and expansive role-playing game (RPG) environm…
Convergence of Muon with Newton-Schulz
Gyu Yeol Kim, Min-hwan Oh
We analyze Muon as originally proposed and used in practice -- using the momentum orthogonalization with a few Newton-Schulz steps. The prior theoretical results replace this key s…
ADAM Optimization with Adaptive Batch Selection
Gyu Yeol Kim, Min-hwan Oh
Adam is a widely used optimizer in neural network training due to its adaptive learning rate. However, because different data samples influence model updates to varying degrees, tr…
A Humanoid Visual-Tactile-Action Dataset for Contact-Rich Manipulation
Eunju Kwon, Seungwon Oh, In-Chang Baek +5
Contact-rich manipulation has become increasingly important in robot learning. However, previous studies on robot learning datasets have focused on rigid objects and underrepresent…