Utilizing Excess Resources in Training Neural Networks
2022-07-12Code Available0· sign in to hype
Amit Henig, Raja Giryes
Code Available — Be the first to reproduce this paper.
ReproduceCode
- github.com/amithenig/kfloOfficialIn paperpytorch★ 0
Abstract
In this work, we suggest Kernel Filtering Linear Overparameterization (KFLO), where a linear cascade of filtering layers is used during training to improve network performance in test time. We implement this cascade in a kernel filtering fashion, which prevents the trained architecture from becoming unnecessarily deeper. This also allows using our approach with almost any network architecture and let combining the filtering layers into a single layer in test time. Thus, our approach does not add computational complexity during inference. We demonstrate the advantage of KFLO on various network models and datasets in supervised learning.