3 papers
cs.SE2026
Studying Quality Improvements Recommended via Manual and Automated Code Review
Giuseppe Crupi, Rosalia Tufano, Gabriele Bavota
Several Deep Learning (DL)-based techniques have been proposed to automate code review. Still, it is unclear the extent to which these approaches can recommend quality improvements…
cs.SE2026
Improving Code Generation via Small Language Model-as-a-judge
Giuseppe Crupi, Rosalia Tufano, Gabriele Bavota
Large language models (LLMs) have shown remarkable capabilities in automated code generation. While effective for mainstream languages, they may underperform on less common or doma…
cs.SE2025
On the Effectiveness of LLM-as-a-judge for Code Generation and Summarization
Giuseppe Crupi, Rosalia Tufano, Alejandro Velasco +3
Large Language Models have been recently exploited as judges for complex natural language processing tasks, such as Q&A. The basic idea is to delegate to an LLM the assessment of t…