Showing cs.AIShow all
3 papers · 1 filter
cs.AI2026
ParBench: A Benchmark for Reliable Evaluation of LLM Parallel Code Translation
Samyak Jhaveri, Erel Kaplan, Tom Yotam +4
Modern compute-intensive software must migrate across a changing ecosystem of accelerators, programming APIs, compiler stacks, and portability layers, including CUDA, OpenMP, OpenC…
cs.AI2026
An Agentic Evaluation Framework for AI-Generated Scientific Code in PETSc
Hong Zhang, Barry Smith, Satish Balay +4
While large language models have significantly accelerated scientific code generation, comprehensively evaluating the generated code remains a major challenge. Traditional benchmar…
cs.AI2025
AI Assistants to Enhance and Exploit the PETSc Knowledge Base
Barry Smith, Junchao Zhang, Hong Zhang +7
Generative AI, especially through large language models (LLMs), is transforming how technical knowledge can be accessed, reused, and extended. PETSc, a widely used numerical librar…