1 citations · 1 across the 3 of their papers we have counts for
1 paper · 1 filter
Shangyu Li, Juyong Jiang, Tiancheng Zhao +1
We introduce OSVBench, a new benchmark for evaluating Large Language Models (LLMs) on the task of generating complete formal specifications for verifying the functional correctness…