Your browser doesn't support javascript.
loading
Trans-Balance: Reducing demographic disparity for prediction models in the presence of class imbalance.
Hong, Chuan; Liu, Molei; Wojdyla, Daniel M; Hickey, Jimmy; Pencina, Michael; Henao, Ricardo.
Afiliação
  • Hong C; Duke University, Department of Biostatistics and Bioinformatics, Durham, NC, USA; Duke Clinical Research Institute, Durham, NC, USA. Electronic address: Chuan.Hong@duke.edu.
  • Liu M; Columbia University, Department of Biostatistics, New York, NY, USA.
  • Wojdyla DM; Duke Clinical Research Institute, Durham, NC, USA.
  • Hickey J; North Carolina State University, Department of Statistics, Raleigh, NC, USA.
  • Pencina M; Duke University, Department of Biostatistics and Bioinformatics, Durham, NC, USA; Duke Clinical Research Institute, Durham, NC, USA.
  • Henao R; Duke University, Department of Biostatistics and Bioinformatics, Durham, NC, USA; Duke Clinical Research Institute, Durham, NC, USA.
J Biomed Inform ; 149: 104532, 2024 Jan.
Article em En | MEDLINE | ID: mdl-38070817
INTRODUCTION: Risk prediction, including early disease detection, prevention, and intervention, is essential to precision medicine. However, systematic bias in risk estimation caused by heterogeneity across different demographic groups can lead to inappropriate or misinformed treatment decisions. In addition, low incidence (class-imbalance) outcomes negatively impact the classification performance of many standard learning algorithms which further exacerbates the racial disparity issues. Therefore, it is crucial to improve the performance of statistical and machine learning models in underrepresented populations in the presence of heavy class imbalance. METHOD: To address demographic disparity in the presence of class imbalance, we develop a novel framework, Trans-Balance, by leveraging recent advances in imbalance learning, transfer learning, and federated learning. We consider a practical setting where data from multiple sites are stored locally under privacy constraints. RESULTS: We show that the proposed Trans-Balance framework improves upon existing approaches by explicitly accounting for heterogeneity across demographic subgroups and cohorts. We demonstrate the feasibility and validity of our methods through numerical experiments and a real application to a multi-cohort study with data from participants of four large, NIH-funded cohorts for stroke risk prediction. CONCLUSION: Our findings indicate that the Trans-Balance approach significantly improves predictive performance, especially in scenarios marked by severe class imbalance and demographic disparity. Given its versatility and effectiveness, Trans-Balance offers a valuable contribution to enhancing risk prediction in biomedical research and related fields.
Assuntos
Palavras-chave

Texto completo: 1 Coleções: 01-internacional Base de dados: MEDLINE Assunto principal: Algoritmos / Pesquisa Biomédica Limite: Humans Idioma: En Revista: J Biomed Inform Assunto da revista: INFORMATICA MEDICA Ano de publicação: 2024 Tipo de documento: Article

Texto completo: 1 Coleções: 01-internacional Base de dados: MEDLINE Assunto principal: Algoritmos / Pesquisa Biomédica Limite: Humans Idioma: En Revista: J Biomed Inform Assunto da revista: INFORMATICA MEDICA Ano de publicação: 2024 Tipo de documento: Article