基于IRT的迫选测验分数信度估计

Reliability Estimates for IRT-Based Forced-Choice Assessment Scores

ORGANIZATIONAL RESEARCH METHODS · 2021
被引 12
人大 A-ABS 4

中文导读

系统比较了基于瑟斯顿IRT模型的迫选测验分数的多种信度估计方法,通过模拟和实证研究揭示其理论差异与数值差异,为实践者选择合适方法提供指南。

Abstract

Forced-choice (FC) assessments of noncognitive psychological constructs (e.g., personality, behavioral tendencies) are popular in high-stakes organizational testing scenarios (e.g., informing hiring decisions) due to their enhanced resistance against response distortions (e.g., faking good, impression management). The measurement precisions of FC assessment scores used to inform personnel decisions are of paramount importance in practice. Different types of reliability estimates are reported for FC assessment scores in current publications, while consensus on best practices appears to be lacking. In order to provide understanding and structure around the reporting of FC reliability, this study systematically examined different types of reliability estimation methods for Thurstonian IRT-based FC assessment scores: their theoretical differences were discussed, and their numerical differences were illustrated through a series of simulations and empirical studies. In doing so, this study provides a practical guide for appraising different reliability estimation methods for IRT-based FC assessment scores.

心理测量学项目反应理论组织心理学人事测评