3 citations · 3 across the 2 of their papers we have counts for
3 papers
Evaluating the Impact of Explainable AI on Trust in AI-Assisted Code Review
Zhenhan Gao, Marvin Muñoz Barón, Umm-e Habiba +2
Background: Large language models (LLMs) are increasingly used to automate code review, but the reasoning behind their decisions remains hard to understand. Developers struggle to…
Guidelines for Empirical Studies in Software Engineering involving Large Language Models
Sebastian Baltes, Florian Angermeir, Chetan Arora +19
Large Language Models (LLMs) are widely used in software engineering (SE) research and practice, yet their non-determinism, opaque training data, and rapidly evolving models threat…
Towards Evaluation Guidelines for Empirical Studies involving LLMs
Stefan Wagner, Marvin Muñoz Barón, Davide Falessi +1
In the short period since the release of ChatGPT, large language models (LLMs) have changed the software engineering research landscape. While there are numerous opportunities to u…