6 citations · 6 across the 2 of their papers we have counts for
Showing cs.SEShow all
2 papers · 1 filter
cs.SE2026
OpenHarmony Bench: Evaluating LLMs and Coding Agents on OpenHarmony App Development
Li Li, Han Hu, Tianjian Zhang +27
We present OPENHARMONY BENCH, an app-level coding benchmark for evaluating LLM-based coding agents on OpenHarmony ArkTS applications. Unlike function-level benchmarks, it evaluates…
cs.SE2024★ 6 cited
TestART: Improving LLM-based Unit Testing via Co-evolution of Automated Generation and Repair Iteration
Siqi Gu, Quanjun Zhang, Kecheng Li +5
Unit testing is crucial for detecting bugs in individual program units but consumes time and effort. Recently, large language models (LLMs) have demonstrated remarkable capabilitie…