12 citations · 14 across the 3 of their papers we have counts for
4 papers
Finding Fast Transformers: One-Shot Neural Architecture Search by Component Composition
Henry Tsai, Jayden Ooi, Chun-Sung Ferng +2
Transformer-based models have achieved stateof-the-art results in many tasks in natural language processing. However, such models are usually slow at inference time, making deploym…
ConQUR: Mitigating Delusional Bias in Deep Q-learning
Andy Su, Jayden Ooi, Tyler Lu +2
Delusional bias is a fundamental source of error in approximate Q-learning. To date, the only techniques that explicitly address delusion require comprehensive search using tabular…
Data Efficient Training for Reinforcement Learning with Adaptive Behavior Policy Sharing
Ge Liu, Rui Wu, Heng-Tze Cheng +7
Deep Reinforcement Learning (RL) is proven powerful for decision making in simulated environments. However, training deep RL model is challenging in real world applications such as…
Advantage Amplification in Slowly Evolving Latent-State Environments
Martin Mladenov, Ofer Meshi, Jayden Ooi +2
Latent-state environments with long horizons, such as those faced by recommender systems, pose significant challenges for reinforcement learning (RL). In this work, we identify and…