Title
From 'F' to 'A' on the N.Y. Regents Science Exams: An Overview of the Aristo Project
Abstract
AI has achieved remarkable mastery over games such as Chess, Go, and Poker, and even Jeopardy, but the rich variety of standardized exams has remained a landmark challenge. Even in 2016, the best AI system achieved merely 59.3% on an 8th Grade science exam challenge. This paper reports unprecedented success on the Grade 8 New York Regents Science Exam, where for the first time a system scores more than 90% on the exam's non-diagram, multiple choice (NDMC) questions. In addition, our Aristo system, building upon the success of recent language models, exceeded 83% on the corresponding Grade 12 Science Exam NDMC questions. The results, on unseen test questions, are robust across different test years and different variations of this kind of test. They demonstrate that modern NLP methods can result in mastery on this task. While not a full solution to general question-answering (the questions are multiple choice, and the domain is restricted to 8th Grade science), it represents a significant milestone for the field.
Year
DOI
Venue
2020
10.1609/aimag.v41i4.5304
AI Mag.
DocType
Volume
Issue
Journal
41
4
Citations 
PageRank 
References 
0
0.34
0
Authors
12
Name
Order
Citations
PageRank
Peter Clark120215.11
Oren Etzioni2101031175.08
Khot Tushar300.34
Bhavana Bharat Dalvi420117.31
Kyle Richardson564.47
Ashish Sabharwal6106370.62
Carissa Schoenick7171.62
Oyvind Tafjord8937.94
Niket Tandon914617.32
Bhakthavatsalam Sumithra1000.34
Groeneveld Dirk1100.34
Guerquin Michal1200.34