{"title":"Variable-length fully Bayesian adaptive testing and associated stopping criteria.","authors":"Luping Niu, Seung W Choi","doi":"10.3758/s13428-026-03148-0","DOIUrl":null,"url":null,"abstract":"<p><p>Many computerized adaptive testing (CAT) systems treat item parameters as if they were known without error, relying on point estimates obtained during item pool calibration. This practice can underestimate uncertainty in ability estimates and affect when a variable-length CAT terminates. A fully Bayesian (FB) CAT algorithm addresses this issue by explicitly incorporating item parameter uncertainty into both ability estimation and item selection. This study investigated the performance of FB CAT in a variable-length setting and compared it with conventional CAT under three stopping rules: a standard error (SE) rule, a change-in- <math><mi>θ</mi></math> (CIT) rule, and a combined CIT+SE rule. Simulation studies were conducted across a range of calibration sample sizes and item pool sizes. Results showed that the FB algorithm generally improved estimation accuracy and produced interval coverage rates closer to nominal levels, especially when the calibration sample size was small. The combined CIT+SE rule reduced unnecessarily long tests that can arise when using only the SE rule, particularly at <math><mi>θ</mi></math> levels for which the remaining item pool provides limited additional information to further reduce SE. Overall, the findings indicate that FB variable-length CAT can enhance uncertainty quantification, and that the combined CIT+SE rule offers a practical balance between measurement precision and testing efficiency.</p>","PeriodicalId":8717,"journal":{"name":"Behavior Research Methods","volume":"58 9","pages":""},"PeriodicalIF":5.0000,"publicationDate":"2026-08-17","publicationTypes":"Journal Article","fieldsOfStudy":null,"isOpenAccess":false,"openAccessPdf":"","citationCount":"0","resultStr":null,"platform":"Semanticscholar","paperid":null,"PeriodicalName":"Behavior Research Methods","FirstCategoryId":"102","ListUrlMain":"https://doi.org/10.3758/s13428-026-03148-0","RegionNum":2,"RegionCategory":"心理学","ArticlePicture":[],"TitleCN":null,"AbstractTextCN":null,"PMCID":null,"EPubDate":"","PubModel":"","JCR":"Q1","JCRName":"PSYCHOLOGY, EXPERIMENTAL","Score":null,"Total":0}
引用次数: 0
Abstract
Many computerized adaptive testing (CAT) systems treat item parameters as if they were known without error, relying on point estimates obtained during item pool calibration. This practice can underestimate uncertainty in ability estimates and affect when a variable-length CAT terminates. A fully Bayesian (FB) CAT algorithm addresses this issue by explicitly incorporating item parameter uncertainty into both ability estimation and item selection. This study investigated the performance of FB CAT in a variable-length setting and compared it with conventional CAT under three stopping rules: a standard error (SE) rule, a change-in- (CIT) rule, and a combined CIT+SE rule. Simulation studies were conducted across a range of calibration sample sizes and item pool sizes. Results showed that the FB algorithm generally improved estimation accuracy and produced interval coverage rates closer to nominal levels, especially when the calibration sample size was small. The combined CIT+SE rule reduced unnecessarily long tests that can arise when using only the SE rule, particularly at levels for which the remaining item pool provides limited additional information to further reduce SE. Overall, the findings indicate that FB variable-length CAT can enhance uncertainty quantification, and that the combined CIT+SE rule offers a practical balance between measurement precision and testing efficiency.
期刊介绍:
Behavior Research Methods publishes articles concerned with the methods, techniques, and instrumentation of research in experimental psychology. The journal focuses particularly on the use of computer technology in psychological research. An annual special issue is devoted to this field.