3 papers
cs.CR2026
Continuous Discovery of Vulnerabilities in LLM Serving Systems with Fuzzing
Yunze Zhao, Yibo Zhao, Yuchen Zhang +2
LLM inference and serving systems have become security-critical infrastructure; however, many of their most concerning failures arise from the serving layer rather than from model…
cs.LG2026
Enabling Performant and Flexible Model-Internal Observability for LLM Inference
Nengneng Yu, Sixian Xiong, Yibo Zhao +2
Today's inference-time workloads increasingly depend on timely access to a model's internal states. We present DMI-Lib, a high-speed deep model inspector that treats internal obser…
cs.DB2025
Approximation-First Timeseries Monitoring Query At Scale
Zeying Zhu, Jonathan Chamberlain, Kenny Wu +2
Timeseries monitoring systems such as Prometheus play a crucial role in gaining observability of the underlying system components. These systems collect timeseries metrics from var…