Skip to main content
Have a personal or library account? Click to login
How to Measure the Reliability of Tasks in Experimental Psychology Cover

How to Measure the Reliability of Tasks in Experimental Psychology

By:   
Open Access
|Aug 2026

Abstract

Calculating the reliability of experimental tasks can be surprisingly difficult using existing tools. Although R packages such as psych are robust, they often require data in wide format and assume carefully selected items that avoid floor and ceiling effects. To encourage the reporting of task reliability in experimental research, I have written an R function, ICC_participants_long, which uses the intraclass correlation coefficient (ICC) to measure the reliability of participant scores directly from data in long format. Applying this function revealed that the current split-half approach may underestimate the reliability of experimental tasks. Furthermore, the model-based approach makes it possible to generate Best Linear Unbiased Predictions (BLUPs) as estimates of participants’ scores, which provides an informative supplement to the raw means.

DOI: https://doi.org/10.5334/joc.516 | Journal eISSN: 2514-4820
Language: English
Page range: 42 - 42
Submitted on: May 3, 2026
Accepted on: Aug 6, 2026
Published on: Aug 18, 2026
Published by: Ubiquity Press
In partnership with: Paradigm Publishing Services

© 2026 Marc Brysbaert, published by Ubiquity Press
This work is licensed under the Creative Commons Attribution 4.0 License.