4 papers
Deployment Risk Assessment Using Diff-Aware Features: A Case Study at Prime Video
Mayur Kurup, Hyunjae Suh, Swathi Vaidyanathan +3
At Amazon Prime Video, we face the critical operational challenge of managing code deployments during live events and rapid feature releases without causing service outages. Curren…
Human or LLM? A Comparative Study on Accessible Code Generation Capability
Hyunjae Suh, Mahan Tafreshipour, Sam Malek +1
Web accessibility is essential for inclusive digital experiences, yet the accessibility of LLM-generated code remains underexplored. This paper presents an empirical study comparin…
An Empirical Study on Automatically Detecting AI-Generated Source Code: How Far Are We?
Hyunjae Suh, Mahan Tafreshipour, Jiawei Li +2
Artificial Intelligence (AI) techniques, especially Large Language Models (LLMs), have started gaining popularity among researchers and software developers for generating source co…
Does the Order of Fine-tuning Matter and Why?
Qihong Chen, Jiawei Li, Hyunjae Suh +5
To improve the performance on a target task, researchers have fine-tuned language models with an intermediate task before the target task of interest. However, previous works have…