3 citations · 4 across the 3 of their papers we have counts for
3 papers
cs.AR2024
Cambricon-LLM: A Chiplet-Based Hybrid Architecture for On-Device Inference of 70B LLM
Zhongkai Yu, Shengwen Liang, Tianyun Ma +12
Deploying advanced large language models on edge devices, such as smartphones and robotics, is a growing trend that enhances user data privacy and network connectivity resilience w…
cs.LG2023★ 1 cited
Online Prototype Alignment for Few-shot Policy Transfer
Qi Yi, Rui Zhang, Shaohui Peng +10
Domain adaptation in reinforcement learning (RL) mainly deals with the changes of observation when transferring the policy to a new environment. Many traditional approaches of doma…
cs.LG2022★ 3 cited
Object-Category Aware Reinforcement Learning
Qi Yi, Rui Zhang, Shaohui Peng +6
Object-oriented reinforcement learning (OORL) is a promising way to improve the sample efficiency and generalization ability over standard RL. Recent works that try to solve OORL t…