2 papers
cs.LG2024
Off-Policy Selection for Initiating Human-Centric Experimental Design
Ge Gao, Xi Yang, Qitong Gao +3
In human-centric tasks such as healthcare and education, the heterogeneity among patients and students necessitates personalized treatments and instructional interventions. While r…
cs.LG2021
InferNet for Delayed Reinforcement Tasks: Addressing the Temporal Credit Assignment Problem
Markel Sanz Ausin, Hamoon Azizsoltani, Song Ju +2
The temporal Credit Assignment Problem (CAP) is a well-known and challenging task in AI. While Reinforcement Learning (RL), especially Deep RL, works well when immediate rewards ar…