171 citations · 280 across the 12 of their papers we have counts for
3 papers · 1 filter
LIT: Block-wise Intermediate Representation Training for Model Compression
Animesh Koratana, Daniel Kang, Peter Bailis +1
Knowledge distillation (KD) is a popular method for reducing the computational overhead of deep network inference, in which the output of a teacher model is used to train a smaller…
Analysis of DAWNBench, a Time-to-Accuracy Machine Learning Performance Benchmark
Cody Coleman, Daniel Kang, Deepak Narayanan +7
Researchers have proposed hardware, software, and algorithmic optimizations to improve the computational performance of deep learning. While some of these optimizations perform the…
BlazeIt: Optimizing Declarative Aggregation and Limit Queries for Neural Network-Based Video Analytics
Daniel Kang, Peter Bailis, Matei Zaharia
Recent advances in neural networks (NNs) have enabled automatic querying of large volumes of video data with high accuracy. While these deep NNs can produce accurate annotations of…