4 papers
Invariant Learning via Probability of Sufficient and Necessary Causes
Mengyue Yang, Zhen Fang, Yonggang Zhang +5
Out-of-distribution (OOD) generalization is indispensable for learning models in the wild, where testing distribution typically unknown and different from the training. Recent meth…
ChessGPT: Bridging Policy Learning and Language Modeling
Xidong Feng, Yicheng Luo, Ziyan Wang +6
When solving decision-making tasks, humans typically depend on information from two key sources: (1) Historical policy data, which provides interaction replay from the environment,…
Interpretable Reward Redistribution in Reinforcement Learning: A Causal Approach
Yudi Zhang, Yali Du, Biwei Huang +4
A major challenge in reinforcement learning is to determine which state-action pairs are responsible for future rewards that are delayed. Reward redistribution serves as a solution…
Linking Health News to Research Literature
Jun Wang, Bei Yu
Accurately linking news articles to scientific research works is a critical component in a number of applications, such as measuring the social impact of a research work and detect…