Fitting Polytomous Item Response Theory Models to Multiple-Choice Tests
- 1 June 1995
- journal article
- Published by SAGE Publications in Applied Psychological Measurement
- Vol. 19 (2) , 143-166
- https://doi.org/10.1177/014662169501900203
Abstract
This study examined how well current software implementations of four polytomous item response theory models fit several multiple-choice tests. The models were Bock's (1972) nominal model, Samejima's (1979) multiple-choice Model C, Thissen & Steinberg's (1984) multiple-choice model, and Levine's (1993) maximum-likelihood formula scoring model. The parameters of the first three of these models were estimated with Thissen's (1986) MULTILOG computer program; Williams & Levine's (1993) FORSCORE program was used for Levine's model. Tests from the Armed Services Vocational Aptitude Battery,the Scholastic Aptitude Test, and the American College Test Assessment were analyzed. The models were fit in estimation samples of approximately 3,000; cross-validation samples of approximately 3,000 were used to evaluate goodness of fit. Both fit plots and X2 statistics were used to determine the adequacy of fit. Bock's model provided surprisingly good fit; adding parameters to the nominal model did not yield improvements in fit. FORSCORE provided generally good fit for Levine's nonparametric model across all tests. Index terms: Bock's nominal model, FORSCORE, maximum likelihood formula scoring, MULTILOG, polytomous IRT.Keywords
This publication has 4 references indexed in Scilit:
- The Number of Guttman Errors as a Simple and Powerful Person-Fit StatisticApplied Psychological Measurement, 1994
- A Nonparametric Approach to the Analysis of Dichotomous Item ResponsesApplied Psychological Measurement, 1982
- The technic of homogeneous tests compared with some aspects of "scale analysis" and factor analysis.Psychological Bulletin, 1948
- A systematic approach to the construction and evaluation of tests of ability.Psychological Monographs: General and Applied, 1947