2 papers
cs.SE2026
Evaluating LLM-Generated Code: A Benchmark and Developer Study
Joanna Szych, Anne Schwerk
Code generation is one of the tasks for which the use of Large Language Models is widely adopted and highly successful. Given this popularity, there are many benchmarks dedicated t…
cs.CL2026
Enhancing Sentiment Classification and Irony Detection in Large Language Models through Advanced Prompt Engineering Techniques
Marvin Schmitt, Anne Schwerk, Sebastian Lempert
This study investigates the use of prompt engineering to enhance large language models (LLMs), specifically GPT-4o-mini and gemini-1.5-flash, in sentiment analysis tasks. It evalua…