2 papers
cs.DC2026
psRL: Efficient Training for Agentic AI via Training-Time Prefix Sharing
Mianjie Yu, Zizhao Mo, Huanyu Qu +8
In modern agentic AI training, the system bottleneck is shifting from rollout to update. Emerging sampling strategies such as tree-structured and step-wise RL greatly increase trai…
cs.LG2024
Online Resource Allocation for Edge Intelligence with Colocated Model Retraining and Inference
Huaiguang Cai, Zhi Zhou, Qianyi Huang
With edge intelligence, AI models are increasingly pushed to the edge to serve ubiquitous users. However, due to the drift of model, data, and task, AI model deployed at the edge s…