копировать удалить добавить публикацию в буфер
Запись сообщества
посмотреть историю данной записи
URL
DOI
BibTeX
EndNote
APA
Chicago
DIN 1505
Harvard
MSOffice XML

Are deep ResNets provably better than linear predictors?

C. Yun, S. Sra, и A. Jadbabaie. (2019)cite arxiv:1907.03922Comment: 17 pages.

Аннотация

Recently, a residual network (ResNet) with a single residual block has been shown to outperform linear predictors, in the sense that all its local minima are at least as good as the best linear predictor. We take a step towards extending this result to deep ResNets. As motivation, we first show that there exist datasets for which all local minima of a fully-connected ReLU network are no better than the best linear predictor, while a ResNet can have strictly better local minima. Second, we show that even at its global minimum, the representation obtained from the residual blocks of a 2-block ResNet does not necessarily improve monotonically as more blocks are added, highlighting a fundamental difficulty in analyzing deep ResNets. Our main result on deep ResNets shows that (under some geometric conditions) any critical point is either (i) at least as good as the best linear predictor; or (ii) the Hessian at this critical point has a strictly negative eigenvalue. Finally, we complement our results by analyzing near-identity regions of deep ResNets, obtaining size-independent upper bounds for the risk attained at critical points as well as the Rademacher complexity.

Описание

[1907.03922] Are deep ResNets provably better than linear predictors?

Линки и ресурсы

ключ BibTeX: yun2019resnets
тип записи: article
год: 2019
url: http://arxiv.org/abs/1907.03922
Примечание: cite arxiv:1907.03922Comment: 17 pages

тэги

@kirk86- тэги данного пользователя выделены

Цитировать эту публикацию

искать в

Метаданные

Последнее изменение 6 лет назад
Создан 6 лет назад

Комментарии и рецензии
(0)

Комментарии, или рецензии отсутствуют. Вы можете их написать!