Rater Variability

"I have a question about the limits of variability in the difficulty or challenge posed by different elements of the facets analyzed in a performance. Let us say that a data set derived from a large-scale performance assessment that has the following characteristics:

3. each examinee is rated on both tasks by a random pair of raters from a pool of 50.

the raters vary in harshness from -2 to +2 logits, with Infit Mean-Square between 0.7 and 1.5

the tasks range in difficulty from -0.5 to +0.5 logits with Infit MnSq between 0.9 and 1.2

the assessment items range in challenge from -1 to +1 logit, Infit MnSq between 0.8 and 1.2

"The biggest problem seems to be rater variability. Can Rasch analysis produce fair ability estimates with these large measure and fit variations?"

In this example, variability in rater severity could be an asset. The range of task and item difficulties is small relative to the examinee range. The wide range of rater severity would cause candidates of the same ability to be evaluated against different levels of the rating scale producing both better examinee measures and better validation of rating scale functioning. As long as the raters are self-consistent (across time and across examinees), I can't imagine how variability in severity would ever be a problem.

The variation in your rater fit statistics indicates that some part of your data may be of doubtful quality. This could be due to raters with different rating styles (e.g., halo effect, extremism). If so, you can discover the amount of mis-measurement this causes by allowing each rater to define their own rating scale. The person measures from this model can then be compared to the shared rating scale model. I have a paper using these methods, "Unmodelled Rater Discrimination Error", given at IOMW, 1998.

Looking at your quality-control fit statistics, your tasks and items are performing well. Raters with noisy fit statistics near 1.5 are problematic. Perhaps the misfitting raters encountered idiosyncratic examinees. Drop idiosyncratic examinees and judges from the data temporarily. Analyze the remaining examinees, items, tasks and raters. Verify that the 6 category rating scale is working as intended for all items, tasks and raters by allowing each of these in turn to have their own rating scale definitions. Finally anchor all measures at their most defensible values and reintroduce dropped examinees and judges for the final measurement report.

Rater Variability Lumley T, Congdon P., Linacre J. … Rasch Measurement Transactions, 1999, 12:4 p.

Rasch Books and Publications
Invariant Measurement: Using Rasch Models in the Social, Behavioral, and Health Sciences, 2nd Edn. George Engelhard, Jr. & Jue Wang	Applying the Rasch Model (Winsteps, Facets) 4th Ed., Bond, Yan, Heene	Advances in Rasch Analyses in the Human Sciences (Winsteps, Facets) 1st Ed., Boone, Staver	Advances in Applications of Rasch Measurement in Science Education, X. Liu & W. J. Boone	Rasch Analysis in the Human Sciences (Winsteps) Boone, Staver, Yale
Introduction to Many-Facet Rasch Measurement (Facets), Thomas Eckes	Statistical Analyses for Language Testers (Facets), Rita Green	Invariant Measurement with Raters and Rating Scales: Rasch Models for Rater-Mediated Assessments (Facets), George Engelhard, Jr. & Stefanie Wind	Aplicação do Modelo de Rasch (Português), de Bond, Trevor G., Fox, Christine M	Appliquer le modèle de Rasch: Défis et pistes de solution (Winsteps) E. Dionne, S. Béland
Exploring Rating Scale Functioning for Survey Research (R, Facets), Stefanie Wind	Rasch Measurement: Applications, Khine	Winsteps Tutorials - free Facets Tutorials - free	Many-Facet Rasch Measurement (Facets) - free, J.M. Linacre	Fairness, Justice and Language Assessment (Winsteps, Facets), McNamara, Knoch, Fan
Other Rasch-Related Resources: Rasch Measurement YouTube Channel
Rasch Measurement Transactions & Rasch Measurement research papers - free	An Introduction to the Rasch Model with Examples in R (eRm, etc.), Debelak, Strobl, Zeigenfuse	Rasch Measurement Theory Analysis in R, Wind, Hua	Applying the Rasch Model in Social Sciences Using R, Lamprianou	El modelo métrico de Rasch: Fundamentación, implementación e interpretación de la medida en ciencias sociales (Spanish Edition), Manuel González-Montesinos M.
Rasch Models: Foundations, Recent Developments, and Applications, Fischer & Molenaar	Probabilistic Models for Some Intelligence and Attainment Tests, Georg Rasch	Rasch Models for Measurement, David Andrich	Constructing Measures, Mark Wilson	Best Test Design - free, Wright & Stone Rating Scale Analysis - free, Wright & Masters
Virtual Standard Setting: Setting Cut Scores, Charalambos Kollias	Diseño de Mejores Pruebas - free, Spanish Best Test Design	A Course in Rasch Measurement Theory, Andrich, Marais	Rasch Models in Health, Christensen, Kreiner, Mesba	Multivariate and Mixture Distribution Rasch Models, von Davier, Carstensen

Go to Institute for Objective Measurement Home Page. The Rasch Measurement SIG (AERA) thanks the Institute for Objective Measurement for inviting the publication of Rasch Measurement Transactions on the Institute's website, www.rasch.org.

Coming Rasch-related Events
Apr. 21 - 22, 2025, Mon.-Tue.	International Objective Measurement Workshop (IOMW) - Boulder, CO, www.iomw.net
Jan. 17 - Feb. 21, 2025, Fri.-Fri.	On-line workshop: Rasch Measurement - Core Topics (E. Smith, Winsteps), www.statistics.com
Feb. - June, 2025	On-line course: Introduction to Classical Test and Rasch Measurement Theories (D. Andrich, I. Marais, RUMM2030), University of Western Australia
Feb. - June, 2025	On-line course: Advanced Course in Rasch Measurement Theory (D. Andrich, I. Marais, RUMM2030), University of Western Australia
May 16 - June 20, 2025, Fri.-Fri.	On-line workshop: Rasch Measurement - Core Topics (E. Smith, Winsteps), www.statistics.com
June 20 - July 18, 2025, Fri.-Fri.	On-line workshop: Rasch Measurement - Further Topics (E. Smith, Facets), www.statistics.com
July 21 - 23, 2025, Mon.-Wed.	Pacific Rim Objective Measurement Symposium (PROMS) 2025, www.proms2025.com
Oct. 3 - Nov. 7, 2025, Fri.-Fri.	On-line workshop: Rasch Measurement - Core Topics (E. Smith, Winsteps), www.statistics.com