Yan-ni Jhan, Wan-cen Li, Shin-hui Ruan, Jia-jyun Sie, Iebin Lian
{"title":"Optimal dichotomization of bimodal Gaussian mixtures","authors":"Yan-ni Jhan, Wan-cen Li, Shin-hui Ruan, Jia-jyun Sie, Iebin Lian","doi":"10.1007/s00362-023-01521-1","DOIUrl":null,"url":null,"abstract":"<p>Despite criticism for loss of information and power, dichotomization of variables is still frequently used in social, behavioral, and medical sciences, mainly because it yields more interpretable conclusions for research outcomes and is useful for decision making. However, the artificial choice of cut-points can be controversial and needs proper justification. In this work, we investigate the properties of point-biserial correlation after dichotomization with underlying bimodal Gaussian mixture distributions. We propose a dichotomous grouping procedure that considers the largest standardized difference in group mean while minimizing information loss.</p>","PeriodicalId":51166,"journal":{"name":"Statistical Papers","volume":"21 1","pages":""},"PeriodicalIF":1.2000,"publicationDate":"2024-01-02","publicationTypes":"Journal Article","fieldsOfStudy":null,"isOpenAccess":false,"openAccessPdf":"","citationCount":"0","resultStr":null,"platform":"Semanticscholar","paperid":null,"PeriodicalName":"Statistical Papers","FirstCategoryId":"100","ListUrlMain":"https://doi.org/10.1007/s00362-023-01521-1","RegionNum":3,"RegionCategory":"数学","ArticlePicture":[],"TitleCN":null,"AbstractTextCN":null,"PMCID":null,"EPubDate":"","PubModel":"","JCR":"Q2","JCRName":"STATISTICS & PROBABILITY","Score":null,"Total":0}
引用次数: 0
Abstract
Despite criticism for loss of information and power, dichotomization of variables is still frequently used in social, behavioral, and medical sciences, mainly because it yields more interpretable conclusions for research outcomes and is useful for decision making. However, the artificial choice of cut-points can be controversial and needs proper justification. In this work, we investigate the properties of point-biserial correlation after dichotomization with underlying bimodal Gaussian mixture distributions. We propose a dichotomous grouping procedure that considers the largest standardized difference in group mean while minimizing information loss.
期刊介绍:
The journal Statistical Papers addresses itself to all persons and organizations that have to deal with statistical methods in their own field of work. It attempts to provide a forum for the presentation and critical assessment of statistical methods, in particular for the discussion of their methodological foundations as well as their potential applications. Methods that have broad applications will be preferred. However, special attention is given to those statistical methods which are relevant to the economic and social sciences. In addition to original research papers, readers will find survey articles, short notes, reports on statistical software, problem section, and book reviews.