Showing cs.LGShow all
2 papers · 1 filter
cs.LG2026
Toward Robust and Efficient ML-Based GPU Caching for Modern Inference
Peng Chen, Jiaji Zhang, Hailiang Zhao +11
In modern GPU inference, cache efficiency remains a major bottleneck, and heuristic policies such as \textsc{LRU} can perform far worse than the offline optimum. Existing learning-…
cs.LG2024
Learning-Augmented Algorithms for the Bahncard Problem
Hailiang Zhao, Xueyan Tang, Peng Chen +1
In this paper, we study learning-augmented algorithms for the Bahncard problem. The Bahncard problem is a generalization of the ski-rental problem, where a traveler needs to irrevo…