2 papers
cs.LG2025
Multi-agent Markov Entanglement
Shuze Chen, Tianyi Peng
Value decomposition has long been a fundamental technique in multi-agent dynamic programming and reinforcement learning (RL). Specifically, the value function of a global state $(s…
stat.ME2025
Improving the Estimation of Lifetime Effects in A/B Testing via Treatment Locality
Shuze Chen, David Simchi-Levi, Chonghuan Wang
Utilizing randomized experiments to evaluate the effect of short-term treatments on the short-term outcomes has been well understood and become the golden standard in industrial pr…