Showing cs.LGShow all
3 papers · 1 filter
cs.LG2023
Invariant Learning via Probability of Sufficient and Necessary Causes
Mengyue Yang, Zhen Fang, Yonggang Zhang +5
Out-of-distribution (OOD) generalization is indispensable for learning models in the wild, where testing distribution typically unknown and different from the training. Recent meth…
cs.LG2023
ChessGPT: Bridging Policy Learning and Language Modeling
Xidong Feng, Yicheng Luo, Ziyan Wang +6
When solving decision-making tasks, humans typically depend on information from two key sources: (1) Historical policy data, which provides interaction replay from the environment,…
cs.LG2023
Interpretable Reward Redistribution in Reinforcement Learning: A Causal Approach
Yudi Zhang, Yali Du, Biwei Huang +4
A major challenge in reinforcement learning is to determine which state-action pairs are responsible for future rewards that are delayed. Reward redistribution serves as a solution…