2 papers
cs.CV2020
Channel-wise Knowledge Distillation for Dense Prediction
Changyong Shu, Yifan Liu, Jianfei Gao +2
Knowledge distillation (KD) has been proven to be a simple and effective tool for training compact models. Almost all KD variants for dense prediction tasks align the student and t…
cs.LG2020
Infinity Learning: Learning Markov Chains from Aggregate Steady-State Observations
Jianfei Gao, Mohamed A. Zahran, Amit Sheoran +2
We consider the task of learning a parametric Continuous Time Markov Chain (CTMC) sequence model without examples of sequences, where the training data consists entirely of aggrega…