arrow
返回

Peano: learning formal mathematical reasoning

delete2023-06-05
delete8
delete
OA
AI
G
Gabriel Poesia *
N
Noah D. Goodman
DOI:10.1098/rsta.2022.0044delete
delete原文链接
delete原文求助
delete分享
delete收藏
摘要

摘要

En 中文
General mathematical reasoning is computationally undecidable, but humans routinely solve new problems. Moreover, discoveries developed over centuries are taught to subsequent generations quickly. What structure enables this, and how might that inform automated mathematical reasoning? We posit that central to both puzzles is the structure of procedural abstractions underlying mathematics. We explore this idea in a case study on five sections of beginning algebra on the Khan Academy platform. To define a computational foundation, we introduce Peano, a theorem-proving environment where the set of valid actions at any point is finite. We use Peano to formalize introductory algebra problems and axioms, obtaining well-defined search problems. We observe existing reinforcement learning methods for symbolic reasoning to be insufficient to solve harder problems. Adding the ability to induce reusable abstractions ('tactics') from its own solutions allows an agent to make steady progress, solving all problems. Furthermore, these abstractions induce an order to the problems, seen at random during training. The recovered order has significant agreement with the expert-designed Khan Academy curriculum, and second-generation agents trained on the recovered curriculum learn significantly faster. These results illustrate the synergistic role of abstractions and curricula in the cultural transmission of mathematics.This article is part of a discussion meeting issue 'Cognitive artificial intelligence'.
Keyword:
automated theorem proving
mathematical reasoning
reinforcement learning
curriculum learning
library learning

期刊

P
Philosophical Transactions of the Royal Society A-Mathematical Physical and Engineering Sciences
IF:
3.7
论文数:
7.8K
被引数:
2.8W

机构

S
Stanford University
学者数:
9.6W
论文数: 8.2W
被引数: 17.0W
引用论文

引用论文

S1-Leitlinie Post-COVID/Long-COVIDS1-Leitlinie后COVID/Long-COVID
err2021-09-02
err0
errOAAI
errAndreas Rembert Koczulla; Tobias Ankermann; Uta Behrends; Peter Berlit; Sebastian Böing; Folke Brinkmann; Christian Franke; Rainer Glöckl; Christian Gogoll; Thomas Hummel; Juliane Kronsbein; Thomas Maibaum; Eva M. J. Peters; Michael Pfeifer; Thomas Platz; Matthias Pletz; Georg Pongratz; Frank Powitz; Klaus F. Rabe; Carmen Scheibenbogen; Andreas Stallmach; Michael Stegbauer; Hans Otto Wagner; Christiane Waller; Hubert Wirtz; Andreas Zeiher; Ralf Harun Zwick
err分享
err收藏
Monoclonal Antibodies Selective for Low Molecular Weight Neurofilaments
err2002-02-01
err0
PREAI
errNiklas Norgren; Jan-Erik Karlsson; Lars Rosengren; Torgny Stigbrand
err分享
err收藏
Traveling Wave Solutions of Parabolic Systems
err1994-10-21
err0
errOAAI
errAizik Volpert; Vitaly Volpert; Vladimir Volpert
err分享
err收藏
err分享
err收藏
err分享
err收藏
学者 查看更多内容