A comparative study of AI-human-made and human-made test forms for a university TESOL theory course

Kyung Mi O

Journal ArticleOPEN ACCESS

A comparative study of AI-human-made and human-made test forms for a university TESOL theory course

Language Testing in Asia (2024) 14(1)

DOI: 10.1186/s40468-024-00291-3

0Citations

56Readers

Abstract

This study examines the efficacy of artificial intelligence (AI) in creating parallel test items compared to human-made ones. Two test forms were developed: one consisting of 20 existing human-made items and another with 20 new items generated with ChatGPT assistance. Expert reviews confirmed the content parallelism of the two test forms. Forty-three university students then completed the 40 test items presented randomly from both forms on a final test. Statistical analyses of student performance indicated comparability between the AI-human-made and human-made test forms. Despite limitations such as sample size and reliance on classical test theory (CTT), the findings suggest ChatGPT’s potential to assist teachers in test item creation, reducing workload and saving time. These results highlight ChatGPT’s value in educational assessment and emphasize the need for further research and development in this area.

Author supplied keywords

Cite

CITATION STYLE

APA

O, K. M. (2024). A comparative study of AI-human-made and human-made test forms for a university TESOL theory course. Language Testing in Asia, 14(1). https://doi.org/10.1186/s40468-024-00291-3

A comparative study of AI-human-made and human-made test forms for a university TESOL theory course

Abstract

Author supplied keywords

Cite

Register to see more suggestions