1 citations · 1 across the 2 of their papers we have counts for
3 papers
cs.CL2024
InverseCoder: Self-improving Instruction-Tuned Code LLMs with Inverse-Instruct
Yutong Wu, Di Huang, Wenxuan Shi +13
Recent advancements in open-source code large language models (LLMs) have been driven by fine-tuning on the data generated from powerful closed-source LLMs, which are expensive to…
cs.LG2023★ 1 cited
Online Prototype Alignment for Few-shot Policy Transfer
Qi Yi, Rui Zhang, Shaohui Peng +10
Domain adaptation in reinforcement learning (RL) mainly deals with the changes of observation when transferring the policy to a new environment. Many traditional approaches of doma…
cs.LG2021
Eden: A Unified Environment Framework for Booming Reinforcement Learning Algorithms
Ruizhi Chen, Xiaoyu Wu, Yansong Pan +12
With AlphaGo defeats top human players, reinforcement learning(RL) algorithms have gradually become the code-base of building stronger artificial intelligence(AI). The RL algorithm…