考试(生物学)
可比性
数学教育
心理学
工作量
试验设计
成就测验
计算机科学
试验方法
统计
数学
标准化测试
生物
操作系统
组合数学
古生物学
标识
DOI:10.1186/s40468-024-00291-3
摘要
Abstract This study examines the efficacy of artificial intelligence (AI) in creating parallel test items compared to human-made ones. Two test forms were developed: one consisting of 20 existing human-made items and another with 20 new items generated with ChatGPT assistance. Expert reviews confirmed the content parallelism of the two test forms. Forty-three university students then completed the 40 test items presented randomly from both forms on a final test. Statistical analyses of student performance indicated comparability between the AI-human-made and human-made test forms. Despite limitations such as sample size and reliance on classical test theory (CTT), the findings suggest ChatGPT’s potential to assist teachers in test item creation, reducing workload and saving time. These results highlight ChatGPT’s value in educational assessment and emphasize the need for further research and development in this area.
科研通智能强力驱动
Strongly Powered by AbleSci AI