2 papers
cs.NI2026
MAESTRO: Multi-Agent Evaluation Suite for Testing, Reliability, and Observability
Tie Ma, Yixi Chen, Vaastav Anand +8
We present MAESTRO, an evaluation suite for the testing, reliability, and observability of LLM-based MAS. MAESTRO standardizes MAS configuration and execution through a unified int…
cs.LG2025
Where is the Testbed for my Federated Learning Research?
Janez BožiÄ, Amândio R. Faustino, Boris RadoviÄ +2
Progressing beyond centralized AI is of paramount importance, yet, distributed AI solutions, in particular various federated learning (FL) algorithms, are often not comprehensively…