PulseExploreJournal ClubDebatesTrendingResearchersJournals
Instagram
HomeExploreJournal ClubTrending
Synapse
⌘+K
Synapse
January 1, 2022SAGE Open40 citationsOpen Access

Artificial Intelligence-Generated and Human Expert-Designed Vocabulary Tests: A Comparative Study

LYLuo YunjiuWWWei WeiYZYing Zheng

Key Points

Key points are not available for this paper at this time.

Abstract

Artificial intelligence (AI) technologies have the potential to reduce the workload for the second language (L2) teachers and test developers. We propose two AI distractor-generating methods for creating Chinese vocabulary items: semantic similarity and visual similarity. Semantic similarity refers to antonyms and synonyms, while visual similarity refers to the phenomenon that two phrases share one or more characters in common. This study explores the construct validity of the two types of selected-response vocabulary tests (AI-generated items and human expert-designed items) and compares their item difficulty and item discrimination. Both quantitative and qualitative data were collected. Seventy-eight students from Beijing Language and Culture University were asked to respond to AI-generated and human expert-designed items respectively. Students’ scores were analyzed using the two-parameter item response theory (2PL-IRT) model. Thirteen students were then invited to report their test taking strategies in the think-aloud section. The findings from the students’ item responses revealed that the human expert-designed items were easier but had more discriminating power than the AI-generated items. The results of think-aloud data indicated that the AI-generated items and expert-designed items might assess different constructs, in which the former elicited test takers’ bottom-up test-taking strategies while the latter seemed more likely to trigger test takers’ rote memorization ability.

Ask AI
Helpful
Bookmark
Share
View Full Paper

Cite This Study

Yunjiu et al. (2022) studied this question.

synapsesocial.com/papers/6a16bfae7cba52b0f77b8e8chttps://doi.org/10.1177/21582440221082130
Ask AI
Helpful
Bookmark
Share
View Full Paper