Applied Psychological Measurement · 2017 · 30 citations · 30 references
EngineeringGeneralizability TheoryData PreparationItem Response TheoryEducationPsychometricsClassical Test TheoryPlausible-value Imputation StatisticsPsychologyLatent ModelingData MiningApplied MeasurementFactor AnalysisLatent Trait EstimatesStatisticsLatent Variable MethodsReliabilityItem FitTest DevelopmentPredictive AnalyticsKnowledge DiscoveryLatent Variable ModelEducational TestingParametric Bootstrap ProcedureEducational MeasurementData TreatmentStatistical InferencePsychological MeasurementFailure Prediction
When tests consist of a small number of items, the use of latent trait estimates for secondary analyses is problematic. One area in particular where latent trait estimates have been problematic is when testing for item misfit. This article explores the use of plausible-value imputations to lessen the severity of the inherent measurement unreliability in shorter tests, and proposes a parametric bootstrap procedure to generate empirical sampling characteristics for null-hypothesis tests of item fit. Simulation results suggest that the proposed item-fit statistics provide conservative to nominal error detection rates. Power to detect item misfit tended to be less than Stone’s [Formula: see text] item-fit statistic but higher than the [Formula: see text] statistic proposed by Orlando and Thissen, especially in tests with 20 or more dichotomously scored items.
30