1 paper
Susana Haing, Natan Vidra, Spurthi Setty
Code generated by LLMs can violate a developer's implicit intentions when given an ambiguous prompt, yet standard benchmarks measure only whether code passes its stated test. We in…