2 papers
cs.DC2025
Counting Without Running: Evaluating LLMs' Reasoning About Code Complexity
Gregory Bolet, Giorgis Georgakoudis, Konstantinos Parasyris +4
Modern GPU software stacks demand developers who can anticipate performance bottlenecks before ever launching a kernel; misjudging floating-point workloads upstream can derail tuni…
cs.DC2025
Can Large Language Models Predict Parallel Code Performance?
Gregory Bolet, Giorgis Georgakoudis, Harshitha Menon +5
Accurate determination of the performance of parallel GPU code typically requires execution-time profiling on target hardware -- an increasingly prohibitive step due to limited acc…