From the 1 of 1 linked paper with an AI index.
1 paper
Cesare Zavattari, Alessandro Tommasi, Giuseppe Prencipe
The paper studies how a single human can audit a large fleet of LLM agents under a limited audit budget, analyzing how miscalibrated confidence scores and correlated errors affect…