1 paper
Shivam Pankaj Kumar, Swati Bararia, Kislay Raj
We present a systematic evaluation of five large language models on automated code review, comparing Claude Sonnet 4.6, Claude Haiku 4.5, GPT-5.4 mini, Minimax M2.7, and GLM-5 Turb…