3 papers
cs.AI2026
Learning When to Optimize: Verified Optimization Skills from Expert GPU-Kernel Lineages
Shuoming Zhang, Qiuchu Yu, Yangyu Zhang +6
LLM-based agents are increasingly used to generate GPU kernels, but they often know what optimizations to try without knowing when those optimizations are sound. We introduce KLine…
cs.CR2026
A Novel Byte-Level Flow-to-Image Encoding Method for Network Intrusion Detection Systems
Ziyu Mu, Zihui Yan, Xiyu Shi +1
Network-based Intrusion Detection Systems (IDS) are predominantly trained on tabular flow records, whose one-dimensional representations limit convolutional architectures from expl…
cs.PL2026
Hexagon-MLIR: An AI Compilation Stack For Qualcomm's Neural Processing Units (NPUs)
Mohammed Javed Absar, Muthu Baskaran, Abhikrant Sharma +22
In this paper, we present Hexagon-MLIR,an open-source compilation stack that targets Qualcomm Hexagon Neural Processing Unit (NPU) and provides unified support for lowering Triton…