2 papers
cs.DC2026
ReviveMoE: Fast Recovery for Hardware Failures in Large-Scale MoE LLM Inference Deployments
Haley Li, Xinglu Wang, Cong Feng +12
As LLM deployments scale over more hardware, the probability of a single failure in a system increases significantly, and cloud operators must consider robust countermeasures to ha…
cs.SE2026
LLM-based Vulnerability Detection at Project Scale: An Empirical Study
Fengjie Li, Jiajun Jiang, Dongchi Chen +1
In this paper, we present the first comprehensive empirical study of specialized LLM-based detectors and compare them with traditional static analyzers at the project scale. Specif…