Audit to Forget: A Unified Method to Revoke Patients' Private Data in Intelligent Healthcare
arXiv:2302.09813 · doi:10.1038/s41467-023-41703-x
Abstract
Revoking personal private data is one of the basic human rights, which has already been sheltered by several privacy-preserving laws in many countries. However, with the development of data science, machine learning and deep learning techniques, this right is usually neglected or violated as more and more patients' data are being collected and used for model training, especially in intelligent healthcare, thus making intelligent healthcare a sector where technology must meet the law, regulations, and privacy principles to ensure that the innovation is for the common good. In order to secure patients' right to be forgotten, we proposed a novel solution by using auditing to guide the forgetting process, where auditing means determining whether a dataset has been used to train the model and forgetting requires the information of a query dataset to be forgotten from the target model. We unified these two tasks by introducing a new approach called knowledge purification. To implement our solution, we developed AFS, a unified open-source software, which is able to evaluate and revoke patients' private data from pre-trained deep learning models. We demonstrated the generality of AFS by applying it to four tasks on different datasets with various data sizes and architectures of deep learning networks. The software is publicly available at \url{https://github.com/JoshuaChou2018/AFS}.
References in corpus (6)
- Distilling the Knowledge in a Neural Network
- MedMNIST v2 -- A large-scale lightweight benchmark for 2D and 3D biomedical image classification
- Fast Yet Effective Machine Unlearning
- A Survey of Machine Unlearning
- Markov Chain Monte Carlo-Based Machine Unlearning: Unlearning What Needs to be Forgotten
- SSSE: Efficiently Erasing Samples from Trained Machine Learning Models