Logo image
The development and use of adapted criterion-referenced reliability indices for composite measures with subscale-level cut scores
Dissertation   Open access

The development and use of adapted criterion-referenced reliability indices for composite measures with subscale-level cut scores

Tingting Chen
University of Iowa
Doctor of Philosophy (PhD), University of Iowa
Spring 2026
DOI: 10.25820/etd.008326
pdf
PhD_Thesis_TingtingChen_FINAL_0430261.98 MBDownloadView
Open Access

Abstract

Decisions made from scores for composite measures with embedded subscales frequently rely on one or more cut scores. Research remains scarce on criterion-referenced reliability specific to subscale-level cut scores. In this study, existing indices for criterion-referenced reliability were extended to accommodate the multivariate situation of a composite measure following a conjunctive decision rule. Two major categories of criterion-referenced reliability were investigated: Squared-error loss indices such as Livingston’s and ϕ(λ) coefficients adapted respectively under classical test theory and generalizability theory, and threshold loss indices including classification consistency and accuracy adapted using a modified multinomial approach. Both original and adapted indices were applied to polytomously scored composite (Conscientiousness) and facet subscale (Organization, Responsibility, and Productiveness) scores from the BFI-2 under the respective conditions of identical and different subscale-level cut scores. Original and adapted Livingston’s and ϕ(λ) coefficients were shown to be equivalent when identical cut scores were used at either composite or subscale levels and exhibited the same pattern of being smallest in value when cut scores were at the mean and became increasingly larger as cut scores deviated farther away from the mean. Original and adapted classification consistency and accuracy showed a similar pattern of variation as those observed in previous research, with consistency indices having an overall inverse relationship with score density and an overall positive association with test length when other conditions were held constant. The differences between conjunctive and compensatory classification indices were also affected by subscale score combinations counting towards the probability of success. Future research is recommended to explore other decision rules using simulation-based analyses with alternative estimation procedures.

Details

Metrics

1 Record Views
Logo image