2 papers
cs.PF2026
KernelSight-LM: A Kernel-Level LLM Inference Simulator
Xiteng Yao, Taeho Kim, Hengzhi Pei +7
As large language models (LLMs) move into production serving, practitioners must rapidly evaluate inference performance across diverse hardware, models, and serving parameters to m…
cs.CE2025
VeBPF Many-Core Architecture for Network Functions in FPGA-based SmartNICs and IoT
Zaid Tahir, Ahmed Sanaullah, Sahan Bandara +2
FPGA-based SmartNICs and IoT devices integrating soft-processors for network function execution have emerged to address the limited hardware reconfigurability of DPUs and MCUs. How…