15 papers
From Incomplete Architecture to Quantified Risk: Multimodal LLM-Driven Security Assessment for Cyber-Physical Systems
Shaofei Huang, Christopher M. Poskitt, Lwin Khin Shar
Cyber-physical systems often contend with incomplete architectural documentation or outdated information resulting from legacy technologies, knowledge management gaps, and the comp…
SafeClaw-R: Towards Safe and Secure Multi-Agent Personal Assistants
Haoyu Wang, Zibo Xiao, Yedi Zhang +2
LLM-based multi-agent systems (MASs) are transforming personal productivity by autonomously executing complex, cross-platform tasks. Frameworks such as OpenClaw demonstrate the pot…
ProbGuard: Proactive Runtime Monitoring for LLM Agent Safety via Probabilistic Prediction
Haoyu Wang, Christopher M. Poskitt, Jiali Wei +1
Large Language Model (LLM) agents increasingly operate across domains such as robotics, virtual assistants, and web automation. However, their stochastic decision-making introduces…
Natural Adversaries: Fuzzing Autonomous Vehicles with Realistic Roadside Object Placements
Yang Sun, Haoyu Wang, Christopher M. Poskitt +1
The emergence of Autonomous Vehicles (AVs) has spurred research into testing the resilience of their perception systems, i.e., ensuring that they are not susceptible to critical mi…
Bayesian and Multi-Objective Decision Support for Real-Time Incident Mitigation in Critical Infrastructure
Shaofei Huang, Christopher M. Poskitt, Lwin Khin Shar
Critical infrastructure increasingly relies on interconnected cyber-physical systems whose security incidents can escalate rapidly into safety and operational failures. Existing de…
Rethinking Artifact Evaluation for Software Engineering in the Age of Generative AI
Christoph Treude, Christopher M. Poskitt, Rashina Hoda
Peer review in software engineering research operates under tight time constraints, while generative AI has substantially reduced the human effort required to produce polished rese…