2 papers
cs.LG2025
Causal Policy Learning in Reinforcement Learning: Backdoor-Adjusted Soft Actor-Critic
Thanh Vinh Vo, Young Lee, Haozhe Ma +2
Hidden confounders that influence both states and actions can bias policy learning in reinforcement learning (RL), leading to suboptimal or non-generalizable behavior. Most RL algo…
cs.LG2025
The Race to Efficiency: A New Perspective on AI Scaling Laws
Chien-Ping Lu
As large-scale AI models expand, training becomes costlier and sustaining progress grows harder. Classical scaling laws (e.g., Kaplan et al. (2020), Hoffmann et al. (2022)) predict…