2 papers
cs.LG2019
Automatic Compiler Based FPGA Accelerator for CNN Training
Shreyas Kolala Venkataramanaiah, Yufei Ma, Shihui Yin +4
Training of convolutional neural networks (CNNs)on embedded platforms to support on-device learning is earning vital importance in recent days. Designing flexible training hard-war…
cs.NE2019
Efficient Network Construction through Structural Plasticity
Xiaocong Du, Zheng Li, Yufei Ma +1
Deep Neural Networks (DNNs) on hardware is facing excessive computation cost due to the massive number of parameters. A typical training pipeline to mitigate over-parameterization…