Showing 2026Show all
2 papers · 1 filter
cs.CL2026
Last Translation Benchmark
Vilém Zouhar, Niyati Bafna, Mukund Choudhary +241
For scientific progress, we need benchmarks that test the limits of state-of-the-art models, and evaluation methods that inform us about failure cases. As models get stronger, stan…
cs.PL2026
Inferring Empirical Sound Resource Bounds via Symbolic Execution and Linear Programming (Extended Version)
Samuel Frontull, Manuel Meitinger, Georg Moser
Existing approaches to resource analysis of programs can be classified into two main paradigms: static analysis and dynamic analysis methods. The former allow for formal guarantees…