From the 1 of 71 linked papers with an AI index.
6 citations · 6 across the 27 of their papers we have counts for
3 papers · 1 filter
CodeArena: A Collective Evaluation Platform for LLM Code Generation
Mingzhe Du, Anh Tuan Luu, Bin Ji +5
Large Language Models (LLMs) have reshaped code generation by synergizing their exceptional comprehension of natural language and programming syntax, thereby substantially boosting…
SoVAR: Building Generalizable Scenarios from Accident Reports for Autonomous Driving Testing
An Guo, Yuan Zhou, Haoxiang Tian +7
Autonomous driving systems (ADSs) have undergone remarkable development and are increasingly employed in safety-critical applications. However, recently reported data on fatal acci…
Mercury: A Code Efficiency Benchmark for Code Large Language Models
Mingzhe Du, Anh Tuan Luu, Bin Ji +2
Amidst the recent strides in evaluating Large Language Models for Code (Code LLMs), existing benchmarks have mainly focused on the functional correctness of generated code, neglect…