1 paper
Manav Singhal, Tushar Aggarwal, Abhijeet Awasthi +2
Existing evaluation benchmarks of language models of code (code LMs) focus almost exclusively on whether the LMs can generate functionally-correct code. In real-world software engi…