3 papers
cs.CL2026
AMDKernelVault: Large-Scale Datasets and Agentic Training for AMD GPU Kernel Optimization
Ji Liu, Saptarshi Majumder, Yiqing Huang +14
We introduce AMDKernelVault, an open HIP and Triton kernel corpus and training framework for recent AMD CDNA GPUs. Existing LLM-based kernel agents are largely CUDA/NVIDIA-centric…
cs.CL2026
AgentKernelArena: Generalization-Aware Benchmarking of GPU Kernel Optimization Agents
Sharareh Younesian, Wenwen Ouyang, Sina Rafati +11
GPU kernel optimization is increasingly critical for efficient deep learning systems, but writing high-performance kernels still requires substantial low-level expertise. Recent AI…
cs.CL2024
Banishing LLM Hallucinations Requires Rethinking Generalization
Johnny Li, Saksham Consul, Eda Zhou +9
Despite their powerful chat, coding, and reasoning abilities, Large Language Models (LLMs) frequently hallucinate. Conventional wisdom suggests that hallucinations are a consequenc…