Pith. sign in

REVIEW 1 cited by

It's Hard for Neural Networks To Learn the Game of Life

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2009.01398 v1 pith:X5QXTHIL submitted 2020-09-03 cs.LG stat.ML

classification cs.LGstat.ML
keywords networkslearnneuralconvergeconvolutionalfunctiongamelife
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

Efforts to improve the learning abilities of neural networks have focused mostly on the role of optimization methods rather than on weight initializations. Recent findings, however, suggest that neural networks rely on lucky random initial weights of subnetworks called "lottery tickets" that converge quickly to a solution. To investigate how weight initializations affect performance, we examine small convolutional networks that are trained to predict n steps of the two-dimensional cellular automaton Conway's Game of Life, the update rules of which can be implemented efficiently in a 2n+1 layer convolutional network. We find that networks of this architecture trained on this task rarely converge. Rather, networks require substantially more parameters to consistently converge. In addition, near-minimal architectures are sensitive to tiny changes in parameters: changing the sign of a single weight can cause the network to fail to learn. Finally, we observe a critical value d_0 such that training minimal networks with examples in which cells are alive with probability d_0 dramatically increases the chance of convergence to a solution. We conclude that training convolutional neural networks to learn the input/output function represented by n steps of Game of Life exhibits many characteristics predicted by the lottery ticket hypothesis, namely, that the size of the networks required to learn this function are often significantly larger than the minimal network required to implement the function.

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 1 Pith paper

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. AutomataGPT: Forecasting and Ruleset Inference for Two-Dimensional Cellular Automata

    cs.LG 2025-06 conditional novelty 6.0 of 10

    A transformer pretrained on 100 cellular automaton rules forecasts unseen rules at 98.5% one-step accuracy and infers new rules with up to 96% functional accuracy.

Pith tools