Norm-based generalisation bounds for multi-class convolutional neural networks

Ledent, Antoine; Mustafa, Waleed; Lei, Yunwen; Kloft, Marius

Computer Science > Machine Learning

arXiv:1905.12430 (cs)

[Submitted on 29 May 2019 (v1), last revised 21 Feb 2021 (this version, v5)]

Title:Norm-based generalisation bounds for multi-class convolutional neural networks

Authors:Antoine Ledent, Waleed Mustafa, Yunwen Lei, Marius Kloft

View PDF

Abstract:We show generalisation error bounds for deep learning with two main improvements over the state of the art. (1) Our bounds have no explicit dependence on the number of classes except for logarithmic factors. This holds even when formulating the bounds in terms of the $L^2$-norm of the weight matrices, where previous bounds exhibit at least a square-root dependence on the number of classes. (2) We adapt the classic Rademacher analysis of DNNs to incorporate weight sharing -- a task of fundamental theoretical importance which was previously attempted only under very restrictive assumptions. In our results, each convolutional filter contributes only once to the bound, regardless of how many times it is applied. Further improvements exploiting pooling and sparse connections are provided. The presented bounds scale as the norms of the parameter matrices, rather than the number of parameters. In particular, contrary to bounds based on parameter counting, they are asymptotically tight (up to log factors) when the weights approach initialisation, making them suitable as a basic ingredient in bounds sensitive to the optimisation procedure. We also show how to adapt the recent technique of loss function augmentation to our situation to replace spectral norms by empirical analogues whilst maintaining the advantages of our approach.

Subjects:	Machine Learning (cs.LG); Machine Learning (stat.ML)
Cite as:	arXiv:1905.12430 [cs.LG]
	(or arXiv:1905.12430v5 [cs.LG] for this version)
	https://doi.org/10.48550/arXiv.1905.12430

Submission history

From: Antoine Ledent [view email]
[v1] Wed, 29 May 2019 13:28:42 UTC (128 KB)
[v2] Wed, 20 Nov 2019 09:25:25 UTC (154 KB)
[v3] Fri, 7 Feb 2020 09:46:48 UTC (167 KB)
[v4] Thu, 8 Oct 2020 19:54:53 UTC (741 KB)
[v5] Sun, 21 Feb 2021 19:39:43 UTC (803 KB)

Computer Science > Machine Learning

Title:Norm-based generalisation bounds for multi-class convolutional neural networks

Submission history

Access Paper:

References & Citations

DBLP - CS Bibliography

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Computer Science > Machine Learning

Title:Norm-based generalisation bounds for multi-class convolutional neural networks

Submission history

Access Paper:

References & Citations

DBLP - CS Bibliography

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators