3 papers
cs.LG2026
SWE-Proof: Can Language Models Resolve Real-World Issues with Machine-Checked Proofs?
George Ma, Benjamin Mikek, Haoyu Li +9
Ensuring the correctness of LLM-generated code is a core challenge for modern software engineering. Benchmarks for agentic code generation check correctness with held-out test suit…
cs.PL2026
Agentic Code Optimization via Compiler-LLM Cooperation
Benjamin Mikek, Danylo Vashchilenko, Bryan Lu +1
Generating performant executables from high level languages is critical to software performance across a wide range of domains. Modern compilers perform this task by passing code t…
cs.LO2023
A Performance Verification Methodology for Resource Allocation Heuristics
Saksham Goel, Benjamin Mikek, Jehad Aly +3
Performance verification is a nascent but promising tool for understanding the performance and limitations of heuristics under realistic assumptions. Bespoke performance verificati…