Estimating Item Wording Effects in Self-Report Measures with Generalizability Theory-Based SEMs: Illustrations Using the Self-Description Questionnaire-III.
Walter P Vispoel, Hyeri Hong, Hyeryung Lee, Tingting Chen
{"title":"Estimating Item Wording Effects in Self-Report Measures with Generalizability Theory-Based SEMs: Illustrations Using the Self-Description Questionnaire-III.","authors":"Walter P Vispoel, Hyeri Hong, Hyeryung Lee, Tingting Chen","doi":"10.1080/00223891.2026.2628589","DOIUrl":null,"url":null,"abstract":"<p><p>Handling item wording effects within Likert-style, self-report questionnaires has long been a challenge when measuring psychological traits. When testing models for such traits, wording effects are commonly addressed by correlating uniquenesses for negatively and positively phrased items or including separate uncorrelated method factors for each effect. However, the magnitude of wording effects is rarely considered in such analyses or distinguished from effects of multiple sources of measurement error. In this article, we demonstrate how generalizability theory-based structural equation model designs are well suited for such purposes using results from all subscales within the Self-Description Questionnaire-III taken by a large sample of college students (<i>n</i> = 1,796) on two occasions. Results emphasized the importance of separating construct, item wording, and measurement error (specific-factor, transient, and random-response) effects for each individual subscale and the effectiveness of generalizability theory-based techniques in doing so. Within the most complete designs, average proportions of explained observed score variance were highest for targeted constructs, followed respectively by random-response error, transient error, specific-factor error, and item wording. We provide code in R for analyzing both generalizability theory and parallel conventional congeneric structural equation models to estimate construct, wording, and measurement error effects using both single- and multiple-occasion designs.</p>","PeriodicalId":16707,"journal":{"name":"Journal of personality assessment","volume":" ","pages":"597-613"},"PeriodicalIF":2.6000,"publicationDate":"2026-09-01","publicationTypes":"Journal Article","fieldsOfStudy":null,"isOpenAccess":false,"openAccessPdf":"","citationCount":"0","resultStr":null,"platform":"Semanticscholar","paperid":null,"PeriodicalName":"Journal of personality assessment","FirstCategoryId":"102","ListUrlMain":"https://doi.org/10.1080/00223891.2026.2628589","RegionNum":3,"RegionCategory":"心理学","ArticlePicture":[],"TitleCN":null,"AbstractTextCN":null,"PMCID":null,"EPubDate":"2026/2/19 0:00:00","PubModel":"Epub","JCR":"Q2","JCRName":"PSYCHOLOGY, CLINICAL","Score":null,"Total":0}
引用次数: 0
Abstract
Handling item wording effects within Likert-style, self-report questionnaires has long been a challenge when measuring psychological traits. When testing models for such traits, wording effects are commonly addressed by correlating uniquenesses for negatively and positively phrased items or including separate uncorrelated method factors for each effect. However, the magnitude of wording effects is rarely considered in such analyses or distinguished from effects of multiple sources of measurement error. In this article, we demonstrate how generalizability theory-based structural equation model designs are well suited for such purposes using results from all subscales within the Self-Description Questionnaire-III taken by a large sample of college students (n = 1,796) on two occasions. Results emphasized the importance of separating construct, item wording, and measurement error (specific-factor, transient, and random-response) effects for each individual subscale and the effectiveness of generalizability theory-based techniques in doing so. Within the most complete designs, average proportions of explained observed score variance were highest for targeted constructs, followed respectively by random-response error, transient error, specific-factor error, and item wording. We provide code in R for analyzing both generalizability theory and parallel conventional congeneric structural equation models to estimate construct, wording, and measurement error effects using both single- and multiple-occasion designs.
期刊介绍:
The Journal of Personality Assessment (JPA) primarily publishes articles dealing with the development, evaluation, refinement, and application of personality assessment methods. Desirable articles address empirical, theoretical, instructional, or professional aspects of using psychological tests, interview data, or the applied clinical assessment process. They also advance the measurement, description, or understanding of personality, psychopathology, and human behavior. JPA is broadly concerned with developing and using personality assessment methods in clinical, counseling, forensic, and health psychology settings; with the assessment process in applied clinical practice; with the assessment of people of all ages and cultures; and with both normal and abnormal personality functioning.