4 citations · 4 across the 2 of their papers we have counts for
2 papers
cs.LG2026
AIConfigurator: Lightning-Fast Configuration Optimization for Multi-Framework LLM Serving
Tianhao Xu, Yiming Liu, Xianglong Lu +18
Optimizing Large Language Model (LLM) inference in production systems is increasingly difficult due to dynamic workloads, stringent latency/throughput targets, and a rapidly expand…
cs.HC2020★ 4 cited
SirenLess: reveal the intention behind news
Xumeng Chen, Leo Yu-Ho Lo, Huamin Qu
News articles tend to be increasingly misleading nowadays, preventing readers from making subjective judgments towards certain events. While some machine learning approaches have b…