1 citations · 1 across the 3 of their papers we have counts for
1 paper · 1 filter
Stephan Wäldchen
Recent progress towards theoretical interpretability guarantees for AI has been made with classifiers that are based on interactive proof systems. A prover selects a certificate fr…