Loss Surfaces, Mode Connectivity, and Fast Ensembling of DNNs

2018-02-27NeurIPS 2018Code Available1· sign in to hype

Timur Garipov, Pavel Izmailov, Dmitrii Podoprikhin, Dmitry Vetrov, Andrew Gordon Wilson

Code Available — Be the first to reproduce this paper.

Code

github.com/timgaripov/dnn-mode-connectivity
OfficialIn paperpytorch★ 0
github.com/g-benton/loss-surface-simplexes
pytorch★ 100
github.com/biomedia-mbzuai/fissionfusion
pytorch★ 7
github.com/tjwhitaker/prune-and-tune-ensembles
pytorch★ 3
github.com/simon-larsson/keras-swa
tf★ 0
github.com/DeBur19/FGE_reproduction_project
pytorch★ 0
github.com/xuyxu/Ensemble-Pytorch
pytorch★ 0
github.com/chandansharma02/Deep_Learning
pytorch★ 0

Abstract

The loss functions of deep neural networks are complex and their geometric properties are not well understood. We show that the optima of these complex loss functions are in fact connected by simple curves over which training and test accuracy are nearly constant. We introduce a training procedure to discover these high-accuracy pathways between modes. Inspired by this new geometric insight, we also propose a new ensembling method entitled Fast Geometric Ensembling (FGE). Using FGE we can train high-performing ensembles in the time required to train a single model. We achieve improved performance compared to the recent state-of-the-art Snapshot Ensembles, on CIFAR-10, CIFAR-100, and ImageNet.

Loss Surfaces, Mode Connectivity, and Fast Ensembling of DNNs

Code

Abstract

Reproductions