ARTIFICIAL INTELLIGENCE IN EDUCATIONAL MEASUREMENT: CHALLENGES AND OPPORTUNITIES

Authors

  • Akharia Paulina Kehinde, Ph.D Department of Guidance and Counselling, Faculty of Education, Ambrose Alli University, Ekpoma

Keywords:

Artificial Intelligence, Educational Measurement, Assessment, Automated Scoring, Computerized Adaptive Testing, Learning Analytics, Validity, Fairness.

Abstract

Artificial Intelligence (AI) has emerged as a transformative technology with significant implications for educational measurement. The integration of AI into assessment processes has expanded the capabilities of traditional measurement practices through innovations such as automated scoring, computerized adaptive testing, automated item generation, learning analytics, predictive assessment, and personalized feedback systems. This paper examines the opportunities and challenges associated with the application of AI in educational measurement. Drawing on contemporary literature, the paper highlights how AI can enhance assessment efficiency, improve measurement precision, support data-driven decision-making, and facilitate individualized learning experiences. Despite these benefits, several concerns persist regarding the validity, reliability, fairness, transparency, data privacy, and security of AI-driven assessment systems. The study further discusses the risks of algorithmic bias, the lack of explainability in complex AI models, and emerging threats to academic integrity arising from generative AI technologies. The paper argues that the successful adoption of AI in educational measurement requires adherence to established psychometric principles, ethical guidelines, and regulatory frameworks. It concludes that while AI offers considerable potential for improving educational assessment, continuous validation, monitoring, interdisciplinary collaboration, and responsible governance are essential to ensure that AI-based measurement systems remain accurate, equitable, transparent, and beneficial to learners and educational stakeholders.

References

Adadi, A., & Berrada, M. (2018). Peeking inside the black box: A survey on explainable artificial intelligence. IEEE Access, 6, 52138–52160.

American Educational Research Association (AERA), American Psychological Association (APA), & National Council on Measurement in Education (NCME). (2014). Standards for educational and psychological testing. AERA.

Baker, R. S., & Hawn, A. (2021). Algorithmic bias in education. International Journal of Artificial Intelligence in Education, 31(4), 1052–1092.

Bengio, Y., Goodfellow, I., & Courville, A. (2016). Deep learning. MIT Press.

Burstein, J., Tetreault, J., & Madnani, N. (2013). The e-rater® automated essay scoring system. In M. D. Shermis & J. Burstein (Eds.), Handbook of automated essay evaluation (pp. 55–67). Routledge.

European Commission. (2022). Ethics guidelines for trustworthy AI. European Union.

Gierl, M. J., & Lai, H. (2013). Using automatic item generation to create multiple-choice test items. Educational Measurement: Issues and Practice, 32(3), 3–17.

Gierl, M. J., & Lai, H. (2013). Using automatic item generation to develop educational assessments. Educational Measurement: Issues and Practice, 32(3), 3–17.

Goodfellow, I., Bengio, Y., & Courville, A. (2016). Deep learning. MIT Press.

Hambleton, R. K., Swaminathan, H., & Rogers, H. J. (1991). Fundamentals of item response theory. Sage.

Ho, A. D. (2024). Artificial intelligence and educational measurement: Opportunities and threats. Journal of Educational and Behavioral Statistics, 49(5), 1–16.

Holmes, W., Bialik, M., & Fadel, C. (2022). Artificial intelligence in education: Promises and implications for teaching and learning. Center for Curriculum Redesign.

Kasneci, E., Sessler, K., Küchemann, S., Bannert, M., Dementieva, D., Fischer, F., & Kasneci, G. (2023). ChatGPT for good? On opportunities and challenges of large language models for education. Learning and Individual Differences, 103, 102274.

LeCun, Y., Bengio, Y., & Hinton, G. (2015). Deep learning. Nature, 521(7553), 436–444. https://doi.org/10.1038/nature14539

Linn, R. L., & Gronlund, N. E. (2000). Measurement and assessment in teaching (8th ed.). Prentice Hall.

Luckin, R., Holmes, W., Griffiths, M., & Forcier, L. B. (2016). Intelligence unleashed: An argument for AI in education. Pearson.

McCarthy, J. (2007). What is artificial intelligence? Stanford University Computer Science Department. http://jmc.stanford.edu/articles/whatisai.html

Mislevy, R. J. (2018). Sociocognitive foundations of educational measurement. Routledge.

Nitko, A. J., & Brookhart, S. M. (2018). Educational assessment of students (8th ed.). Pearson.

Russell, S., & Norvig, P. (2021). Artificial intelligence: A modern approach (4th ed.). Pearson.

Siemens, G., & Baker, R. S. (2012). Learning analytics and educational data mining: Towards communication and collaboration. Proceedings of the 2nd International Conference on Learning Analytics and Knowledge, 252–254.

UNESCO. (2023). Guidance for generative AI in education and research. UNESCO Publishing.

Williamson, D. M., Xi, X., & Breyer, F. J. (2020). A framework for evaluation and use of automated scoring. Educational Measurement: Issues and Practice, 39(1), 4–16.

Downloads

Published

2026-06-20