Learning Ability of Interpolating Deep Convolutional Neural Networks

Huo, Xiaoming; Zhou, Tian-Yi

Learning Ability of Interpolating Deep Convolutional Neural Networks

Authors: Xiaoming Huo
Tian-Yi Zhou
Publication date: 16 August 2023
Publisher

Abstract

It is frequently observed that overparameterized neural networks generalize well. Regarding such phenomena, existing theoretical work mainly devotes to linear settings or fully-connected neural networks. This paper studies the learning ability of an important family of deep neural networks, deep convolutional neural networks (DCNNs), under both underparameterized and overparameterized settings. We establish the first learning rates of underparameterized DCNNs without parameter or function variable structure restrictions presented in the literature. We also show that by adding well-defined layers to a non-interpolating DCNN, we can obtain some interpolating DCNNs that maintain the good learning rates of the non-interpolating DCNN. This result is achieved by a novel network deepening scheme designed for DCNNs. Our work provides theoretical verification of how overfitted DCNNs generalize well

Similar works

Full text

Available Versions

arXiv.org e-Print Archive

oai:arXiv.org:2210.14184

Last time updated on 04/12/2022