A comparative study of AI-human-made and human-made test forms for a university TESOL theory course

Kyung-Mi O

doi:10.1186/s40468-024-00291-3

Language Testing in Asia (Jun 2024)

A comparative study of AI-human-made and human-made test forms for a university TESOL theory course

Kyung-Mi O

Affiliations

Kyung-Mi O: Dongduk Women’s University

DOI: https://doi.org/10.1186/s40468-024-00291-3
Journal volume & issue: Vol. 14, no. 1
pp. 1 – 17

Abstract

Read online

Abstract This study examines the efficacy of artificial intelligence (AI) in creating parallel test items compared to human-made ones. Two test forms were developed: one consisting of 20 existing human-made items and another with 20 new items generated with ChatGPT assistance. Expert reviews confirmed the content parallelism of the two test forms. Forty-three university students then completed the 40 test items presented randomly from both forms on a final test. Statistical analyses of student performance indicated comparability between the AI-human-made and human-made test forms. Despite limitations such as sample size and reliance on classical test theory (CTT), the findings suggest ChatGPT’s potential to assist teachers in test item creation, reducing workload and saving time. These results highlight ChatGPT’s value in educational assessment and emphasize the need for further research and development in this area.

Published in Language Testing in Asia

ISSN: 2229-0443 (Online)
Publisher: SpringerOpen
Country of publisher: United Kingdom
LCC subjects: Language and Literature
Website: https://languagetestingasia.springeropen.com/

About the journal

Abstract

Keywords