Automated estimation of item difficulty for multiple-choice tests: An application of word embedding techniques. Issue 6 (November 2018)
- Record Type:
- Journal Article
- Title:
- Automated estimation of item difficulty for multiple-choice tests: An application of word embedding techniques. Issue 6 (November 2018)
- Main Title:
- Automated estimation of item difficulty for multiple-choice tests: An application of word embedding techniques
- Authors:
- Hsu, Fu-Yuan
Lee, Hahn-Ming
Chang, Tao-Hsing
Sung, Yao-Ting - Abstract:
- Abstract: Pretesting is the most commonly used method for estimating test item difficulty because it provides highly accurate results that can be applied to assessment development activities. However, pretesting is inefficient, and it can lead to item exposure. Hence, an increasing number of studies have invested considerable effort in researching the automated estimation of item difficulty. Language proficiency tests constitute the majority of researched test topics, while comparatively less research has focused on content subjects. This paper introduces a novel method for the automated estimation of item difficulty for social studies tests. In this study, we explore the difficulty of multiple-choice items, which consist of the following item elements: a question and alternative options. We use learning materials to construct a semantic space using word embedding techniques and project an item's texts into the semantic space to obtain corresponding vectors. Semantic features are obtained by calculating the cosine similarity between the vectors of item elements. Subsequently, these semantic features are sent to a classifier for training and testing. Based on the output of the classifier, an estimation model is created and item difficulty is estimated. Our findings suggest that the semantic similarity between a stem and the options has the strongest impact on item difficulty. Furthermore, the results indicate that the proposed estimation method outperforms pretesting, andAbstract: Pretesting is the most commonly used method for estimating test item difficulty because it provides highly accurate results that can be applied to assessment development activities. However, pretesting is inefficient, and it can lead to item exposure. Hence, an increasing number of studies have invested considerable effort in researching the automated estimation of item difficulty. Language proficiency tests constitute the majority of researched test topics, while comparatively less research has focused on content subjects. This paper introduces a novel method for the automated estimation of item difficulty for social studies tests. In this study, we explore the difficulty of multiple-choice items, which consist of the following item elements: a question and alternative options. We use learning materials to construct a semantic space using word embedding techniques and project an item's texts into the semantic space to obtain corresponding vectors. Semantic features are obtained by calculating the cosine similarity between the vectors of item elements. Subsequently, these semantic features are sent to a classifier for training and testing. Based on the output of the classifier, an estimation model is created and item difficulty is estimated. Our findings suggest that the semantic similarity between a stem and the options has the strongest impact on item difficulty. Furthermore, the results indicate that the proposed estimation method outperforms pretesting, and therefore, we expect that the proposed approach will complement and partially replace pretesting in future. … (more)
- Is Part Of:
- Information processing & management. Volume 54:Issue 6(2018:Nov.)
- Journal:
- Information processing & management
- Issue:
- Volume 54:Issue 6(2018:Nov.)
- Issue Display:
- Volume 54, Issue 6 (2018)
- Year:
- 2018
- Volume:
- 54
- Issue:
- 6
- Issue Sort Value:
- 2018-0054-0006-0000
- Page Start:
- 969
- Page End:
- 984
- Publication Date:
- 2018-11
- Subjects:
- Multiple-choice item -- Item difficulty estimation -- Cognitive processing model -- Semantic similarity -- Word embedding -- Machine learning
Information storage and retrieval systems -- Periodicals
Information science -- Periodicals
Systèmes d'information -- Périodiques
Sciences de l'information -- Périodiques
Information science
Information storage and retrieval systems
Periodicals
658.4038 - Journal URLs:
- http://www.sciencedirect.com/science/journal/03064573 ↗
http://www.elsevier.com/journals ↗ - DOI:
- 10.1016/j.ipm.2018.06.007 ↗
- Languages:
- English
- ISSNs:
- 0306-4573
- Deposit Type:
- Legaldeposit
- View Content:
- Available online (eLD content is only available in our Reading Rooms) ↗
- Physical Locations:
- British Library DSC - 4493.893000
British Library DSC - BLDSS-3PM
British Library HMNTS - ELD Digital store - Ingest File:
- 7213.xml