1 citations · 1 across the 8 of their papers we have counts for
4 papers · 1 filter
Online Scheduling for LLM Inference with KV Cache Constraints
Patrick Jaillet, Jiashuo Jiang, Konstantina Mellou +3
Large Language Model (LLM) inference, where a trained model generates text one word at a time in response to user prompts, is a computationally intensive process requiring efficien…
A Lyapunov Drift-Plus-Penalty Method Tailored for Reinforcement Learning with Queue Stability
Wenhan Xu, Jiashuo Jiang, Lei Deng +1
With the proliferation of Internet of Things (IoT) devices, the demand for addressing complex optimization challenges has intensified. The Lyapunov Drift-Plus-Penalty algorithm is…
Online Pricing and Allocation with Demand Learning and Fulfillment Cost
Jianyu Xu, Xuan Wang, Yu-Xiang Wang +1
We study online learning for a seller that jointly chooses per-period inventory positions and a uniform price, then fulfills realized demand through a downstream allocation. The ma…
Reinforcement Learning with Intrinsically Motivated Feedback Graph for Lost-sales Inventory Control
Zifan Liu, Xinran Li, Shibo Chen +3
Reinforcement learning (RL) has proven to be well-performed and general-purpose in the inventory control (IC). However, further improvement of RL algorithms in the IC domain is imp…