6 citations · 7 across the 4 of their papers we have counts for
Showing cs.SEShow all
2 papers · 1 filter
cs.SE2025
BigCodeArena: Unveiling More Reliable Human Preferences in Code Generation via Execution
Terry Yue Zhuo, Xiaolong Jin, Hange Liu +37
Crowdsourced model evaluation platforms, such as Chatbot Arena, enable real-time evaluation from human perspectives to assess the quality of model responses. In the coding domain,…
cs.SE2023
Automated Code generation for Information Technology Tasks in YAML through Large Language Models
Saurabh Pujar, Luca Buratti, Xiaojie Guo +8
The recent improvement in code generation capabilities due to the use of large language models has mainly benefited general purpose programming languages. Domain specific languages…