3 papers
cs.SE2025
SimpleDevQA: Benchmarking Large Language Models on Development Knowledge QA
Jing Zhang, Lianghong Guo, Yanlin Wang +7
The Development Knowledge Question Answering (Dev Knowledge QA) task aims to provide natural language answers to knowledge-seeking questions during software development. To investi…
cs.SE2025
LLMAID: Identifying AI Capabilities in Android Apps with LLMs
Pei Liu, Terry Zhuo, Jiawei Deng +7
Recent advancements in artificial intelligence (AI) and its widespread integration into mobile software applications have received significant attention, highlighting the growing p…
cs.SE2025
LLM-as-a-Judge for Software Engineering: Literature Review, Vision, and the Road Ahead
Junda He, Jieke Shi, Terry Yue Zhuo +5
The rapid integration of Large Language Models (LLMs) into software engineering (SE) has revolutionized tasks like code generation, producing a massive volume of software artifacts…