activity
20242026
collaborators

7 papers

cs.AI2026

Unaccountable Delegation, Fading Skills: Mapping the Risks of Workplace AI Agents

Gabriele La Malfa, Lakmal Meegahapola, Edyta Bogucka +4

To anticipate socio-technical risks from AI agents, organizations need taxonomies to classify them. However, existing AI risk taxonomies focus on broad risks and do not capture job…

cs.AI2026

Explore, Map, Remember, Decide: Are Embodied VLMs Ready for Safety-Critical Scenarios?

Gabriele La Malfa, Nitay Alon, Emanuele La Malfa +2

Theory of Space framework (ToS) assesses the spatial understanding of curiosity-driven Vision-Language Models (VLMs) under partial observability. As AI techniques are increasingly…

cs.AI2026

The Attacker in the Mirror: Breaking Self-Consistency in Safety via Anchored Bipolicy Self-Play

Gabriele La Malfa, Emanuele La Malfa, Saar Cohen +4

Self-play red team is an established approach to improving AI safety in which different instances of the same model play attacker and defender roles in a zero-sum game, i.e., where…

cs.MA2025

Large Language Models Miss the Multi-Agent Mark

Emanuele La Malfa, Gabriele La Malfa, Samuele Marro +5

Recent interest in Multi-Agent Systems of Large Language Models (MAS LLMs) has led to an increase in frameworks leveraging multiple LLMs to tackle complex tasks. However, much of t…

cs.MA2025

Fairness Aware Reinforcement Learning via Proximal Policy Optimization

Gabriele La Malfa, Jie M. Zhang, Michael Luck +1

Fairness in multi-agent systems (MAS) focuses on equitable reward distribution among agents in scenarios involving sensitive attributes such as race, gender, or socioeconomic statu…

astro-ph.HE2025

Tracing the ejecta structure of SN 1987A: Insights and diagnostics from 3D MHD simulations

S. Orlando, M. Miceli, M. Ono +14

Supernova (SN) 1987A provides a unique window into the aftermath of a massive stellar explosion, offering key insights into the ejecta's morphology, composition, explosion mechanis…