Manuel Plisson, Antoine Moll, Valentine Sarrazin, Denis Charles, Thibault Antoine, Razvan Ionescu, Odile Koehren, Eric Raymond
{"title":"癌症风险的机器学习和创新算法包容性承保方法。","authors":"Manuel Plisson, Antoine Moll, Valentine Sarrazin, Denis Charles, Thibault Antoine, Razvan Ionescu, Odile Koehren, Eric Raymond","doi":"10.17849/insm-50-1-36-48.1","DOIUrl":null,"url":null,"abstract":"<p><strong>Introduction: </strong>-Due to early detection and improved therapies, the prevalence of long-term breast cancer survivors is increasing. This has increased the need for more inclusive underwriting in individuals with a history of breast cancer. Herein, we developed a method using algorithm aiming facilitating the underwriting of multiple parameters in breast cancer survivors.</p><p><strong>Methods: </strong>-Variables and data were extracted from the SEER database and analyzed using 4 different machine learning based algorithms (Logistic Regression, GA2M, Random Forest, and XGBoost) that were compared with Kaplan Meier survival estimates. The performances of these algorithms have been compared with multiple metrics (Log Loss, AUC, and SMR). In situ (non-invasive) and metastatic breast cancer were excluded from this analysis.</p><p><strong>Results: </strong>-Parameters included the pathological subtype, pTNM staging (T: tumor size, N; number of nodes; M presence or absence of metastases), Scarff-Bloom-Richardson grading, the expression of estrogen and progesterone hormone receptors were selected to predict the individual outcome at any time point from diagnosis. While all models had identical performance in terms of statistical metrics (AUC, Log Loss, and SMR), the logistic regression was the one and only model that respects all business constraints and was intelligible for medical and underwriting users.</p><p><strong>Conclusion: </strong>-This study provides insight to develop algorithms to set underwriter-friendly calculators for more accurate risk estimations that can be used to rationalize insurance pricing for breast cancer survivors. This study supports the development of a more inclusive underwriting based on models that can encompass the heterogeneity of several malignancies such as breast cancer.</p>","PeriodicalId":39345,"journal":{"name":"Journal of insurance medicine (New York, N.Y.)","volume":null,"pages":null},"PeriodicalIF":0.0000,"publicationDate":"2023-07-01","publicationTypes":"Journal Article","fieldsOfStudy":null,"isOpenAccess":false,"openAccessPdf":"","citationCount":"0","resultStr":"{\"title\":\"Methods for Inclusive Underwriting of Breast Cancer Risk with Machine Learning and Innovative Algorithms.\",\"authors\":\"Manuel Plisson, Antoine Moll, Valentine Sarrazin, Denis Charles, Thibault Antoine, Razvan Ionescu, Odile Koehren, Eric Raymond\",\"doi\":\"10.17849/insm-50-1-36-48.1\",\"DOIUrl\":null,\"url\":null,\"abstract\":\"<p><strong>Introduction: </strong>-Due to early detection and improved therapies, the prevalence of long-term breast cancer survivors is increasing. This has increased the need for more inclusive underwriting in individuals with a history of breast cancer. Herein, we developed a method using algorithm aiming facilitating the underwriting of multiple parameters in breast cancer survivors.</p><p><strong>Methods: </strong>-Variables and data were extracted from the SEER database and analyzed using 4 different machine learning based algorithms (Logistic Regression, GA2M, Random Forest, and XGBoost) that were compared with Kaplan Meier survival estimates. The performances of these algorithms have been compared with multiple metrics (Log Loss, AUC, and SMR). In situ (non-invasive) and metastatic breast cancer were excluded from this analysis.</p><p><strong>Results: </strong>-Parameters included the pathological subtype, pTNM staging (T: tumor size, N; number of nodes; M presence or absence of metastases), Scarff-Bloom-Richardson grading, the expression of estrogen and progesterone hormone receptors were selected to predict the individual outcome at any time point from diagnosis. While all models had identical performance in terms of statistical metrics (AUC, Log Loss, and SMR), the logistic regression was the one and only model that respects all business constraints and was intelligible for medical and underwriting users.</p><p><strong>Conclusion: </strong>-This study provides insight to develop algorithms to set underwriter-friendly calculators for more accurate risk estimations that can be used to rationalize insurance pricing for breast cancer survivors. This study supports the development of a more inclusive underwriting based on models that can encompass the heterogeneity of several malignancies such as breast cancer.</p>\",\"PeriodicalId\":39345,\"journal\":{\"name\":\"Journal of insurance medicine (New York, N.Y.)\",\"volume\":null,\"pages\":null},\"PeriodicalIF\":0.0000,\"publicationDate\":\"2023-07-01\",\"publicationTypes\":\"Journal Article\",\"fieldsOfStudy\":null,\"isOpenAccess\":false,\"openAccessPdf\":\"\",\"citationCount\":\"0\",\"resultStr\":null,\"platform\":\"Semanticscholar\",\"paperid\":null,\"PeriodicalName\":\"Journal of insurance medicine (New York, N.Y.)\",\"FirstCategoryId\":\"1085\",\"ListUrlMain\":\"https://doi.org/10.17849/insm-50-1-36-48.1\",\"RegionNum\":0,\"RegionCategory\":null,\"ArticlePicture\":[],\"TitleCN\":null,\"AbstractTextCN\":null,\"PMCID\":null,\"EPubDate\":\"\",\"PubModel\":\"\",\"JCR\":\"Q3\",\"JCRName\":\"Medicine\",\"Score\":null,\"Total\":0}","platform":"Semanticscholar","paperid":null,"PeriodicalName":"Journal of insurance medicine (New York, N.Y.)","FirstCategoryId":"1085","ListUrlMain":"https://doi.org/10.17849/insm-50-1-36-48.1","RegionNum":0,"RegionCategory":null,"ArticlePicture":[],"TitleCN":null,"AbstractTextCN":null,"PMCID":null,"EPubDate":"","PubModel":"","JCR":"Q3","JCRName":"Medicine","Score":null,"Total":0}
Methods for Inclusive Underwriting of Breast Cancer Risk with Machine Learning and Innovative Algorithms.
Introduction: -Due to early detection and improved therapies, the prevalence of long-term breast cancer survivors is increasing. This has increased the need for more inclusive underwriting in individuals with a history of breast cancer. Herein, we developed a method using algorithm aiming facilitating the underwriting of multiple parameters in breast cancer survivors.
Methods: -Variables and data were extracted from the SEER database and analyzed using 4 different machine learning based algorithms (Logistic Regression, GA2M, Random Forest, and XGBoost) that were compared with Kaplan Meier survival estimates. The performances of these algorithms have been compared with multiple metrics (Log Loss, AUC, and SMR). In situ (non-invasive) and metastatic breast cancer were excluded from this analysis.
Results: -Parameters included the pathological subtype, pTNM staging (T: tumor size, N; number of nodes; M presence or absence of metastases), Scarff-Bloom-Richardson grading, the expression of estrogen and progesterone hormone receptors were selected to predict the individual outcome at any time point from diagnosis. While all models had identical performance in terms of statistical metrics (AUC, Log Loss, and SMR), the logistic regression was the one and only model that respects all business constraints and was intelligible for medical and underwriting users.
Conclusion: -This study provides insight to develop algorithms to set underwriter-friendly calculators for more accurate risk estimations that can be used to rationalize insurance pricing for breast cancer survivors. This study supports the development of a more inclusive underwriting based on models that can encompass the heterogeneity of several malignancies such as breast cancer.
期刊介绍:
The Journal of Insurance Medicine is a peer reviewed scientific journal sponsored by the American Academy of Insurance Medicine, and is published quarterly. Subscriptions to the Journal of Insurance Medicine are included in your AAIM membership.