PANDA: Facilitating Usable AI Development
arXiv:1804.09997
Abstract
Recent advances in artificial intelligence (AI) and machine learning have created a general perception that AI could be used to solve complex problems, and in some situations over-hyped as a tool that can be so easily used. Unfortunately, the barrier to realization of mass adoption of AI on various business domains is too high because most domain experts have no background in AI. Developing AI applications involves multiple phases, namely data preparation, application modeling, and product deployment. The effort of AI research has been spent mostly on new AI models (in the model training stage) to improve the performance of benchmark tasks such as image recognition. Many other factors such as usability, efficiency and security of AI have not been well addressed, and therefore form a barrier to democratizing AI. Further, for many real world applications such as healthcare and autonomous driving, learning via huge amounts of possibility exploration is not feasible since humans are involved. In many complex applications such as healthcare, subject matter experts (e.g. Clinicians) are the ones who appreciate the importance of features that affect health, and their knowledge together with existing knowledge bases are critical to the end results. In this paper, we take a new perspective on developing AI solutions, and present a solution for making AI usable. We hope that this resolution will enable all subject matter experts (eg. Clinicians) to exploit AI like data scientists.
References in corpus (19)
- Very Deep Convolutional Networks for Large-Scale Image Recognition
- A Survey on Deep Learning in Medical Image Analysis
- SqueezeNet: AlexNet-level accuracy with 50x fewer parameters and <0.5MB model size
- Practical Bayesian Optimization of Machine Learning Algorithms
- Understanding Black-box Predictions via Influence Functions
- A Survey of Model Compression and Acceleration for Deep Neural Networks
- Tensorizing Neural Networks
- Large Batch Training of Convolutional Networks
- CryptoDL: Deep Neural Networks over Encrypted Data
- A Berkeley View of Systems Challenges for AI
- Database Meets Deep Learning: Challenges and Opportunities
- Crypto-Nets: Neural Networks over Encrypted Data
- Variance Reduction in SGD by Distributed Importance Sampling
- Omnivore: An Optimizer for Multi-device Deep Learning on CPUs and GPUs
- Infrastructure for Usable Machine Learning: The Stanford DAWN Project
- Active Sampler: Light-weight Accelerator for Complex Data Analytics at Scale
- UStore: A Distributed Storage With Rich Semantics
- Weld: Rethinking the Interface Between Data-Intensive Applications
- Consistent Bounded-Asynchronous Parameter Servers for Distributed ML