1 paper
Hong Xie, Haoran Gu, Yanying Huang +2
This paper proposes a variant of multiple-play stochastic bandits tailored to resource allocation problems arising from LLM applications, edge intelligence, etc. The model is compo…