Comparison of machine learning models to predict the risk of breast cancer-related lymphedema among breast cancer survivors: a cross-sectional study in China

Jiali Du; Jing Yang; Qing Yang; Xin Zhang; Ling Yuan; Bing Fu

doi:10.3389/fonc.2024.1334082

Frontiers in Oncology (Feb 2024)

Comparison of machine learning models to predict the risk of breast cancer-related lymphedema among breast cancer survivors: a cross-sectional study in China

Jiali Du,
Jing Yang,
Qing Yang,
Xin Zhang,
Ling Yuan,
Bing Fu

Affiliations

Jiali Du: Department of Breast Surgery, Sichuan Clinical Research Center for Cancer, Sichuan Cancer Hospital & Institute, Sichuan Cancer Center, Affiliated Cancer Hospital of University of Electronic Science and Technology of China, Chengdu, China
Jing Yang: Department of Breast Surgery, Sichuan Clinical Research Center for Cancer, Sichuan Cancer Hospital & Institute, Sichuan Cancer Center, Affiliated Cancer Hospital of University of Electronic Science and Technology of China, Chengdu, China
Qing Yang: Department of Nursing, Sichuan Clinical Research Center for Cancer, Sichuan Cancer Hospital & Institute, Sichuan Cancer Center, Affiliated Cancer Hospital of University of Electronic Science and Technology of China, Chengdu, China
Xin Zhang: Department of Breast Surgery, Sichuan Clinical Research Center for Cancer, Sichuan Cancer Hospital & Institute, Sichuan Cancer Center, Affiliated Cancer Hospital of University of Electronic Science and Technology of China, Chengdu, China
Ling Yuan: Department of Breast Surgery, Sichuan Clinical Research Center for Cancer, Sichuan Cancer Hospital & Institute, Sichuan Cancer Center, Affiliated Cancer Hospital of University of Electronic Science and Technology of China, Chengdu, China
Bing Fu: Department of Breast Surgery, Sichuan Clinical Research Center for Cancer, Sichuan Cancer Hospital & Institute, Sichuan Cancer Center, Affiliated Cancer Hospital of University of Electronic Science and Technology of China, Chengdu, China

DOI: https://doi.org/10.3389/fonc.2024.1334082
Journal volume & issue: Vol. 14

Abstract

Read online

ObjectiveThe aim of this study was to develop and validate a series of breast cancer-related lymphoedema risk prediction models using machine learning algorithms for early identification of high-risk individuals to reduce the incidence of postoperative breast cancer lymphoedema.MethodsThis was a retrospective study conducted from January 2012 to July 2022 in a tertiary oncology hospital. Subsequent to the collection of clinical data, variables with predictive capacity for breast cancer-related lymphoedema (BCRL) were subjected to scrutiny utilizing the Least Absolute Shrinkage and Selection Operator (LASSO) technique. The entire dataset underwent a randomized partition into training and test subsets, adhering to a 7:3 distribution. Nine classification models were developed, and the model performance was evaluated based on accuracy, sensitivity, specificity, recall, precision, F-score, and area under curve (AUC) of the ROC curve. Ultimately, the selection of the optimal model hinged upon the AUC value. Grid search and 10-fold cross-validation was used to determine the best parameter setting for each algorithm.ResultsA total of 670 patients were investigated, of which 469 were in the modeling group and 201 in the validation group. A total of 174 had BCRL (25.97%). The LASSO regression model screened for the 13 features most valuable in predicting BCRL. The range of each metric in the test set for the nine models was, in order: accuracy (0.75–0.84), sensitivity (0.50–0.79), specificity (0.79–0.93), recall (0.50–0.79), precision (0.51–0.70), F score (0.56–0.69), and AUC value (0.71–0.87). Overall, LR achieved the best performance in terms of accuracy (0.81), precision (0.60), sensitivity (0.79), specificity (0.82), recall (0.79), F-score (0.68), and AUC value (0.87) for predicting BCRL.ConclusionThe study established that the constructed logistic regression (LR) model exhibits a more favorable amalgamation of accuracy, sensitivity, specificity, recall, and AUC value. This configuration adeptly discerns patients who are at an elevated risk of BCRL. Consequently, this precise identification equips nurses with the means to undertake timely and tailored interventions, thus averting the onset of BCRL.

Published in Frontiers in Oncology

ISSN: 2234-943X (Online)
Publisher: Frontiers Media S.A.
Country of publisher: Switzerland
LCC subjects: Medicine: Internal medicine: Neoplasms. Tumors. Oncology. Including cancer and carcinogens
Website: https://www.frontiersin.org/journals/oncology/

About the journal

Abstract

Keywords