8 papers
The Creation and Analysis of Government AI Transparency Statements in Australia
Shidong Pan, Haochen Gong, Boming Xia +3
Governments increasingly deploy AI in public services, making transparency essential for accountability and public trust. Australia's Standard for AI Transparency Statements (AITS)…
Harnessing Agent Skills: Architectural Patterns and a Reference Architecture for Skill-Mediated LLM Agents
Boming Xia, Liming Zhu, Zhenchang Xing +3
Agent skills externalise reusable agent-facing behavioural knowledge and guidance as persistent artefacts that can be discovered, activated, and interpreted by LLM agents. Although…
Uncertainty Propagation in LLM-Based Systems
Boming Xia, Liming Zhu, Erdun Gao +3
Uncertainty in large language model (LLM)-based systems is often studied at the level of a single model output, yet deployed LLM applications are compound systems in which uncertai…
OntoMetric: An Ontology-Driven LLM-Assisted Framework for Automated ESG Metric Knowledge Graph Generation
Mingqin Yu, Fethi Rabhi, Boming Xia +3
Environmental, Social, and Governance (ESG) metric knowledge is inherently structured, connecting industries, reporting frameworks, metric categories, metrics, and calculation mode…
Improving Methodologies for LLM Evaluations Across Global Languages
Akriti Vij, Benjamin Chua, Darshini Ramiah +43
As frontier AI models are deployed globally, it is essential that their behaviour remains safe and reliable across diverse linguistic and cultural contexts. To examine how current…
Evaluation-Driven Development and Operations of LLM Agents: A Process Model and Reference Architecture
Boming Xia, Qinghua Lu, Liming Zhu +3
Large Language Models (LLMs) have enabled the emergence of LLM agents, systems capable of pursuing under-specified goals and adapting after deployment. Evaluating such agents is ch…