2 papers
cs.LG2026
PAWN: Piece Value Analysis with Neural Networks
Ethan Tang, Hasan Davulcu, Jia Zou +1
Predicting the relative value of any given chess piece in a position remains an open challenge, as a piece's contribution depends on its spatial relationships with every other piec…
cs.LG2026
Offline-Online Reinforcement Learning for Linear Mixture MDPs
Zhongjun Zhang, Sean R. Sinclair
We study offline-online reinforcement learning in linear mixture Markov decision processes (MDPs) under environment shift. In the offline phase, data are collected by an unknown be…