2 papers
cs.SE2025
Who is Introducing the Failure? Automatically Attributing Failures of Multi-Agent Systems via Spectrum Analysis
Yu Ge, Linna Xie, Zhong Li +2
Large Language Model Powered Multi-Agent Systems (MASs) are increasingly employed to automate complex real-world problems, such as programming and scientific discovery. Despite the…
cs.DC2025
MultiKernelBench: A Multi-Platform Benchmark for Kernel Generation
Zhongzhen Wen, Yinghui Zhang, Zhong Li +3
The automatic generation of deep learning (DL) kernels using large language models (LLMs) has emerged as a promising approach to reduce the manual effort and hardware-specific expe…