A copy of this work was available on the public web and has been preserved in the Wayback Machine. The capture dates from 2019; you can also visit the original URL.
The file type is application/pdf
.
The Limitations of Standardized Science Tests as Benchmarks for Artificial Intelligence Research: Position Paper
[article]
2015
arXiv
pre-print
In this position paper, I argue that standardized tests for elementary science such as SAT or Regents tests are not very good benchmarks for measuring the progress of artificial intelligence systems in understanding basic science. The primary problem is that these tests are designed to test aspects of knowledge and ability that are challenging for people; the aspects that are challenging for AI systems are very different. In particular, standardized tests do not test knowledge that is obvious
arXiv:1411.1629v2
fatcat:yqgncd7cvngu3fyojgrg2yniq4