1 paper
Bilal Al-Ahmad, M. Harshvardhan, Khaled El-Fakih +1
Context: Large language models (LLMs) can generate unit tests quickly, but high structural coverage does not establish that those tests execute reliably or detect faults. Existing…