works on

From the 1 of 20 linked papers with an AI index.

collaborators

20 papers

cs.SE2026

Hype Meets Reality: Large Language Models as Mutators in Search-based Automated Program Repair of Simulink-Stateflow Models

Ayesha Irshad, Pablo Valle, Jon Ayerdi +1

Search-based Automated Program Repair (APR) techniques rely on carefully designed mutation operators to explore the space of candidate fixes. Recent advances in Large Language Mode…

cs.SE2026

Delta Debugging for Cyber-Physical Systems with Flaky Test Executions

Pablo Valle, Shaukat Ali, Aitor Arrieta

The paper introduces three delta debugging algorithms that use statistical analysis and repeated executions to isolate minimal failure-inducing inputs for stochastic cyber‑physical…

cs.AI2026

RAG-TESTER: Automated End-to-End Testing of Retrieval-Augmented Large Language Models

Ange Maiztegi, Jon Ayerdi, Miren Illarramendi +1

Retrieval-Augmented Generation (RAG) enables Large Language Models (LLMs) to use external and domain-specific knowledge, but its reliability depends on the interaction between the…

cs.SE2026

Foundation Models for Software Engineering of Cyber-Physical Systems: the Road Ahead

Chengjie Lu, Pablo Valle, Jiahui Wu +4

Foundation Models (FMs), particularly Large Language Models (LLMs), are increasingly used to support various software engineering activities (e.g., coding and testing). Their adopt…

cs.SE2026

MANGO: Automated Multi-Agent Test Oracle Generation for Vision-Language-Action Models

Pablo Valle, Shaukat Ali, Aitor Arrieta +1

Vision-Language-Action (VLA) models are emerging robotic control systems that integrate perception, language understanding, and action generation in a unified architecture. Existin…

cs.SE2026

VISOR: A Vision-Language Model-based Test Oracle for Testing Robots

Prasun Saurabh, Pablo Valle, Aitor Arrieta +2

Testing robots requires assessing whether they perform their intended tasks correctly, dependably, and with high quality, a challenge known as the test oracle problem in software t…