1 paper
Zhihong Liu, Long Qian, Zeyang Liu +3
Decision Transformer (DT) can learn effective policy from offline datasets by converting the offline reinforcement learning (RL) into a supervised sequence modeling task, where the…