Cargando…

Methods for Estimating Item-Score Reliability

Reliability is usually estimated for a test score, but it can also be estimated for item scores. Item-score reliability can be useful to assess the item’s contribution to the test score’s reliability, for identifying unreliable scores in aberrant item-score patterns in person-fit analysis, and for s...

Descripción completa

Detalles Bibliográficos
Autores principales:	Zijlmans, Eva A. O., van der Ark, L. Andries, Tijmstra, Jesper, Sijtsma, Klaas
Formato:	Online Artículo Texto
Lenguaje:	English
Publicado:	SAGE Publications 2018
Materias:	Articles
Acceso en línea:	https://www.ncbi.nlm.nih.gov/pmc/articles/PMC6140096/ https://www.ncbi.nlm.nih.gov/pubmed/30237646 http://dx.doi.org/10.1177/0146621618758290

Descripción
Sumario:	Reliability is usually estimated for a test score, but it can also be estimated for item scores. Item-score reliability can be useful to assess the item’s contribution to the test score’s reliability, for identifying unreliable scores in aberrant item-score patterns in person-fit analysis, and for selecting the most reliable item from a test to use as a single-item measure. Four methods were discussed for estimating item-score reliability: the Molenaar–Sijtsma method (method MS), Guttman’s method [Formula: see text] , the latent class reliability coefficient (method LCRC), and the correction for attenuation (method CA). A simulation study was used to compare the methods with respect to median bias, variability (interquartile range [IQR]), and percentage of outliers. The simulation study consisted of six conditions: standard, polytomous items, unequal [Formula: see text] parameters, two-dimensional data, long test, and small sample size. Methods MS and CA were the most accurate. Method LCRC showed almost unbiased results, but large variability. Method [Formula: see text] consistently underestimated item-score reliabilty, but showed a smaller IQR than the other methods.

Methods for Estimating Item-Score Reliability

Ejemplares similares