copy delete add this publication to your clipboard
community post
history of this post
URL
DOI
BibTeX
EndNote
APA
Chicago
DIN 1505
Harvard
MSOffice XML

Can You Learn an Algorithm? Generalizing from Easy to Hard Problems with Recurrent Networks

A. Schwarzschild, E. Borgnia, A. Gupta, F. Huang, U. Vishkin, M. Goldblum, and T. Goldstein. (2021)cite arxiv:2106.04537.

Abstract

Deep neural networks are powerful machines for visual pattern recognition, but reasoning tasks that are easy for humans may still be difficult for neural models. Humans possess the ability to extrapolate reasoning strategies learned on simple problems to solve harder examples, often by thinking for longer. For example, a person who has learned to solve small mazes can easily extend the very same search techniques to solve much larger mazes by spending more time. In computers, this behavior is often achieved through the use of algorithms, which scale to arbitrarily hard problem instances at the cost of more computation. In contrast, the sequential computing budget of feed-forward neural networks is limited by their depth, and networks trained on simple problems have no way of extending their reasoning to accommodate harder problems. In this work, we show that recurrent networks trained to solve simple problems with few recurrent steps can indeed solve much more complex problems simply by performing additional recurrences during inference. We demonstrate this algorithmic behavior of recurrent networks on prefix sum computation, mazes, and chess. In all three domains, networks trained on simple problem instances are able to extend their reasoning abilities at test time simply by "thinking for longer."

Description

Can You Learn an Algorithm? Generalizing from Easy to Hard Problems with Recurrent Networks

Links and resources

BibTeX key: schwarzschild2021learn
entry type: misc
year: 2021
url: http://arxiv.org/abs/2106.04537
note: cite arxiv:2106.04537

@szhang104's tags highlighted

cognition

Cite this publication

@misc{schwarzschild2021learn, abstract = {Deep neural networks are powerful machines for visual pattern recognition, but reasoning tasks that are easy for humans may still be difficult for neural models. Humans possess the ability to extrapolate reasoning strategies learned on simple problems to solve harder examples, often by thinking for longer. For example, a person who has learned to solve small mazes can easily extend the very same search techniques to solve much larger mazes by spending more time. In computers, this behavior is often achieved through the use of algorithms, which scale to arbitrarily hard problem instances at the cost of more computation. In contrast, the sequential computing budget of feed-forward neural networks is limited by their depth, and networks trained on simple problems have no way of extending their reasoning to accommodate harder problems. In this work, we show that recurrent networks trained to solve simple problems with few recurrent steps can indeed solve much more complex problems simply by performing additional recurrences during inference. We demonstrate this algorithmic behavior of recurrent networks on prefix sum computation, mazes, and chess. In all three domains, networks trained on simple problem instances are able to extend their reasoning abilities at test time simply by "thinking for longer."}, added-at = {2021-07-09T20:25:44.000+0200}, author = {Schwarzschild, Avi and Borgnia, Eitan and Gupta, Arjun and Huang, Furong and Vishkin, Uzi and Goldblum, Micah and Goldstein, Tom}, biburl = {https://www.bibsonomy.org/bibtex/26ef2a4f0b3fa60b6110bb4af02bae7dd/szhang104}, description = {Can You Learn an Algorithm? Generalizing from Easy to Hard Problems with Recurrent Networks}, interhash = {f88546327ff0d7a9a44510041d082a96}, intrahash = {6ef2a4f0b3fa60b6110bb4af02bae7dd}, keywords = {cognition}, note = {cite arxiv:2106.04537}, timestamp = {2021-07-09T20:25:44.000+0200}, title = {Can You Learn an Algorithm? Generalizing from Easy to Hard Problems with Recurrent Networks}, url = {http://arxiv.org/abs/2106.04537}, year = 2021 }

BibSonomy

copy delete add this publication to your clipboard
community post
history of this post
URL
DOI
BibTeX
EndNote
APA
Chicago
DIN 1505
Harvard
MSOffice XML

Can You Learn an Algorithm? Generalizing from Easy to Hard Problems with Recurrent Networks

Abstract

Description

Links and resources

Tags

community

Cite this publication

More citation styles

search on

Meta data

Comments and Reviews
(0)

BibSonomy

copydeleteadd this publication to your clipboardcommunity posthistory of this postURLDOIBibTeXEndNoteAPAChicagoDIN 1505HarvardMSOffice XML Can You Learn an Algorithm? Generalizing from Easy to Hard Problems with Recurrent Networks

Abstract

Description

Links and resources

Tags

community

Cite this publication

More citation styles

search on

Meta data

Comments and Reviews (0)

copy delete add this publication to your clipboard
community post
history of this post
URL
DOI
BibTeX
EndNote
APA
Chicago
DIN 1505
Harvard
MSOffice XML

Can You Learn an Algorithm? Generalizing from Easy to Hard Problems with Recurrent Networks

Comments and Reviews
(0)