Norm Loss: An efficient yet effective regularization method for deep neural networks
Proceedings - International Conference on Pattern Recognition/Proceedings/International Conference on Pattern Recognition, pp. 8812–8818
Abstract
Convolutional neural network training can suffer from diverse issues like exploding or vanishing gradients, scaling-based weight space symmetry and covariant-shift. In order to address these issues, researchers develop weight regularization methods and activation normalization methods. In this work we propose a weight soft-regularization method based on the Oblique manifold. The proposed method uses a loss function which pushes each weight vector to have a norm close to one, i.e. the weight matrix is smoothly steered toward the so-called Oblique manifold. We evaluate our method on the very popular CIFAR-10, CIFAR-100 and ImageNet 2012 datasets using two state-of-the-art architectures, namely the ResNet and wide-ResNet. Our method introduces negligible computational overhead and the results show that it is competitive to the state-of-the-art and in some cases superior to it. Additionally, the results are less sensitive to hyperparameter settings such as batch size and regularization factor.
Authors 4
-
Affiliation as printed
Honda Research Institute Europe GmbH, Offenbach, Germany
-
Thomas Bäck Aachen
Affiliation as printed
Leiden University Leiden Institute of Advanced Computer Science, Leiden, the Netherlands
-
Wei Chen Aachen
Affiliation as printed
Leiden University Leiden Institute of Advanced Computer Science, Leiden, the Netherlands
-
Michael S. Lew Aachen
Affiliation as printed
Leiden University Leiden Institute of Advanced Computer Science, Leiden, the Netherlands
Cited by 3 stored of 3
3 results
No patents citing this paper on Lens.org (checked 2026-10-11).