копировать удалить добавить публикацию в буфер
Запись сообщества
посмотреть историю данной записи
URL
DOI
BibTeX
EndNote
APA
Chicago
DIN 1505
Harvard
MSOffice XML

A Deep Conditioning Treatment of Neural Networks

N. Agarwal, P. Awasthi, и S. Kale. (2020)cite arxiv:2002.01523.

Аннотация

We study the role of depth in training randomly initialized overparameterized neural networks. We give the first general result showing that depth improves trainability of neural networks by improving the conditioning of certain kernel matrices of the input data. This result holds for arbitrary non-linear activation functions, and we provide a characterization of the improvement in conditioning as a function of the degree of non-linearity and the depth of the network. We provide versions of the result that hold for training just the top layer of the neural network, as well as for training all layers, via the neural tangent kernel. As applications of these general results, we provide a generalization of the results of Das et al. (2019) showing that learnability of deep random neural networks with arbitrary non-linear activations (under mild assumptions) degrades exponentially with depth. Additionally, we show how benign overfitting can occur in deep neural networks via the results of Bartlett et al. (2019b).

Описание

[2002.01523] A Deep Conditioning Treatment of Neural Networks

Линки и ресурсы

ключ BibTeX: agarwal2020conditioning
тип записи: article
год: 2020
url: http://arxiv.org/abs/2002.01523
Примечание: cite arxiv:2002.01523

тэги

@kirk86- тэги данного пользователя выделены

Цитировать эту публикацию

искать в

Метаданные

Последнее изменение 4 лет назад
Создан 4 лет назад

Комментарии и рецензии
(0)

Комментарии, или рецензии отсутствуют. Вы можете их написать!