9 papers
Evaluation-Driven Development and Operations of LLM Agents: A Process Model and Reference Architecture
Boming Xia, Qinghua Lu, Liming Zhu +3
Large Language Models (LLMs) have enabled the emergence of LLM agents, systems capable of pursuing under-specified goals and adapting after deployment. Evaluating such agents is ch…
AgentArcEval: An Architecture Evaluation Method for Foundation Model based Agents
Qinghua Lu, Dehai Zhao, Yue Liu +6
The emergence of foundation models (FMs) has enabled the development of highly capable and autonomous agents, unlocking new application opportunities across a wide range of domains…
Bridging Solidity Evolution Gaps: An LLM-Enhanced Approach for Smart Contract Compilation Error Resolution
Likai Ye, Mengliang Li, Dehai Zhao +2
Solidity, the dominant smart contract language for Ethereum, has rapidly evolved with frequent version updates to enhance security, functionality, and developer experience. However…
SHIELDA: Structured Handling of Exceptions in LLM-Driven Agentic Workflows
Jingwen Zhou, Jieshan Chen, Qinghua Lu +2
Large Language Model (LLM) agentic systems are software systems powered by LLMs that autonomously reason, plan, and execute multi-step workflows to achieve human goals, rather than…
When Prompt Engineering Meets Software Engineering: CNL-P as Natural and Robust "APIs'' for Human-AI Interaction
Zhenchang Xing, Yang Liu, Zhuo Cheng +4
With the growing capabilities of large language models (LLMs), they are increasingly applied in areas like intelligent customer service, code generation, and knowledge management.…
Think Like an Engineer: A Neuro-Symbolic Collaboration Agent for Generative Software Requirements Elicitation and Self-Review
Sai Zhang, Zhenchang Xing, Jieshan Chen +5
The vision of End-User Software Engineering (EUSE) is to empower non-professional users with full control over the software development lifecycle. It aims to enable users to drive…