activity
20242026
most citedEvaluation-Driven Development and Operations of LLM Agents: A Process Model and Reference Architecture

2 citations · 2 across the 6 of their papers we have counts for

collaborators
Showing cs.SEShow all

5 papers · 1 filter

cs.SE2026

Uncertainty Propagation in LLM-Based Systems

Boming Xia, Liming Zhu, Erdun Gao +3

Uncertainty in large language model (LLM)-based systems is often studied at the level of a single model output, yet deployed LLM applications are compound systems in which uncertai…

cs.SE20252 cited

Evaluation-Driven Development and Operations of LLM Agents: A Process Model and Reference Architecture

Boming Xia, Qinghua Lu, Liming Zhu +3

Large Language Models (LLMs) have enabled the emergence of LLM agents, systems capable of pursuing under-specified goals and adapting after deployment. Evaluating such agents is ch…

cs.SE2024

Privacy and Copyright Protection in Generative AI: A Lifecycle Perspective

Dawen Zhang, Boming Xia, Yue Liu +6

The advent of Generative AI has marked a significant milestone in artificial intelligence, demonstrating remarkable capabilities in generating realistic images, texts, and data pat…

cs.SE2024

Blockchain-Enabled Accountability in Data Supply Chain: A Data Bill of Materials Approach

Yue Liu, Dawen Zhang, Boming Xia +4

In the era of advanced artificial intelligence, highlighted by large-scale generative models like GPT-4, ensuring the traceability, verifiability, and reproducibility of datasets t…

cs.SE2024

An AI System Evaluation Framework for Advancing AI Safety: Terminology, Taxonomy, Lifecycle Mapping

Boming Xia, Qinghua Lu, Liming Zhu +1

The advent of advanced AI underscores the urgent need for comprehensive safety evaluations, necessitating collaboration across communities (i.e., AI, software engineering, and gove…