Pith. sign in

REVIEW

Reinforcement Learning of Theorem Proving

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 1805.07563 v1 pith:GU5RAH4L submitted 2018-05-19 cs.AI cs.LGcs.LO

classification cs.AIcs.LGcs.LO
keywords learningproblemsreinforcementdomainguidinglargemathematicalproof
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

We introduce a theorem proving algorithm that uses practically no domain heuristics for guiding its connection-style proof search. Instead, it runs many Monte-Carlo simulations guided by reinforcement learning from previous proof attempts. We produce several versions of the prover, parameterized by different learning and guiding algorithms. The strongest version of the system is trained on a large corpus of mathematical problems and evaluated on previously unseen problems. The trained system solves within the same number of inferences over 40% more problems than a baseline prover, which is an unusually high improvement in this hard AI domain. To our knowledge this is the first time reinforcement learning has been convincingly applied to solving general mathematical problems on a large scale.

Discussion (0). Sign in to comment.

Pith tools