A Dataset for Learning University STEM Courses at Scale and Generating Questions at a Human Level
DOI:
https://doi.org/10.1609/aaai.v37i13.27091Keywords:
AI For Education, STEM Courses, Natural Language ProcessingAbstract
We present a new dataset for learning to solve, explain, and generate university-level STEM questions from 27 courses across a dozen departments in seven universities. We scale up previous approaches to questions from courses in the departments of Mechanical Engineering, Materials Science and Engineering, Chemistry, Electrical Engineering, Computer Science, Physics, Earth Atmospheric and Planetary Sciences, Economics, Mathematics, Biological Engineering, Data Systems, and Society, and Statistics. We visualize similarities and differences between questions across courses. We demonstrate that a large foundation model is able to generate questions that are as appropriate and at the same difficulty level as human-written questions.Downloads
Published
2023-09-06
How to Cite
Drori, I., Zhang, S., Chin, Z., Shuttleworth, R., Lu, A., Chen, L., Birbo, B., He, M., Lantigua, P., Tran, S., Hunter, G., Feng, B., Cheng, N., Wang, R., Hicke, Y., Surbehera, S., Raghavan, A., Siemenn, A., Singh, N., Lynch, J., Shporer, A., Verma, N., Buonassisi, T., & Solar-Lezama, A. (2023). A Dataset for Learning University STEM Courses at Scale and Generating Questions at a Human Level. Proceedings of the AAAI Conference on Artificial Intelligence, 37(13), 15921-15929. https://doi.org/10.1609/aaai.v37i13.27091
Issue
Section
EAAI Symposium: AI for Education