9 citations · 9 across the 3 of their papers we have counts for
4 papers
OpenHarmony Bench: Evaluating LLMs and Coding Agents on OpenHarmony App Development
Li Li, Han Hu, Tianjian Zhang +27
We present OPENHARMONY BENCH, an app-level coding benchmark for evaluating LLM-based coding agents on OpenHarmony ArkTS applications. Unlike function-level benchmarks, it evaluates…
Distinguishing LLM-generated from Human-written Code by Contrastive Learning
Xiaodan Xu, Chao Ni, Xinrong Guo +4
Large language models (LLMs), such as ChatGPT released by OpenAI, have attracted significant attention from both industry and academia due to their demonstrated ability to generate…
Investigating White-Box Attacks for On-Device Models
Mingyi Zhou, Xiang Gao, Jing Wu +3
Numerous mobile apps have leveraged deep learning capabilities. However, on-device models are vulnerable to attacks as they can be easily extracted from their corresponding mobile…
Practical Program Repair via Preference-based Ensemble Strategy
Wenkang Zhong, Chuanyi Li, Kui Liu +5
To date, over 40 Automated Program Repair (APR) tools have been designed with varying bug-fixing strategies, which have been demonstrated to have complementary performance in terms…