Can You Learn an Algorithm? Generalizing from Easy to Hard Problems with Recurrent Networks

Schwarzschild, Avi; Borgnia, Eitan; Gupta, Arjun; Huang, Furong; Vishkin, Uzi; Goldblum, Micah; Goldstein, Tom

Computer Science > Machine Learning

arXiv:2106.04537 (cs)

[Submitted on 8 Jun 2021 (v1), last revised 2 Nov 2021 (this version, v2)]

Title:Can You Learn an Algorithm? Generalizing from Easy to Hard Problems with Recurrent Networks

Authors:Avi Schwarzschild, Eitan Borgnia, Arjun Gupta, Furong Huang, Uzi Vishkin, Micah Goldblum, Tom Goldstein

View PDF

Abstract:Deep neural networks are powerful machines for visual pattern recognition, but reasoning tasks that are easy for humans may still be difficult for neural models. Humans possess the ability to extrapolate reasoning strategies learned on simple problems to solve harder examples, often by thinking for longer. For example, a person who has learned to solve small mazes can easily extend the very same search techniques to solve much larger mazes by spending more time. In computers, this behavior is often achieved through the use of algorithms, which scale to arbitrarily hard problem instances at the cost of more computation. In contrast, the sequential computing budget of feed-forward neural networks is limited by their depth, and networks trained on simple problems have no way of extending their reasoning to accommodate harder problems. In this work, we show that recurrent networks trained to solve simple problems with few recurrent steps can indeed solve much more complex problems simply by performing additional recurrences during inference. We demonstrate this algorithmic behavior of recurrent networks on prefix sum computation, mazes, and chess. In all three domains, networks trained on simple problem instances are able to extend their reasoning abilities at test time simply by "thinking for longer."

Subjects:	Machine Learning (cs.LG); Artificial Intelligence (cs.AI)
Cite as:	arXiv:2106.04537 [cs.LG]
	(or arXiv:2106.04537v2 [cs.LG] for this version)
	https://doi.org/10.48550/arXiv.2106.04537

Submission history

From: Avi Schwarzschild [view email]
[v1] Tue, 8 Jun 2021 17:19:48 UTC (421 KB)
[v2] Tue, 2 Nov 2021 12:32:07 UTC (440 KB)

Computer Science > Machine Learning

Title:Can You Learn an Algorithm? Generalizing from Easy to Hard Problems with Recurrent Networks

Submission history

Access Paper:

References & Citations

DBLP - CS Bibliography

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Computer Science > Machine Learning

Title:Can You Learn an Algorithm? Generalizing from Easy to Hard Problems with Recurrent Networks

Submission history

Access Paper:

References & Citations

DBLP - CS Bibliography

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators