Preoperative differentiation of borderline and malignant ovarian tumors using interpretable machine learning

Ovarian cancer remains one of the most lethal gynecological malignancies, with borderline ovarian tumors (BOTs) representing lesions of low malignant potential requiring distinct surgical management. Accurate preoperative differentiation between BOTs and epithelial ovarian cancer (OC) is essential f...

Description complète

Enregistré dans:
Détails bibliographiques
Auteurs principaux: SamadiAfshar, Saber (Auteur) , Azizi, Hossein (Auteur) , Masoudi, Mahla (Auteur) , SamadiAfshar, Sahel (Auteur) , Nikakhtar, Ali (Auteur) , Skutella, Thomas (Auteur)
Format: Article (Journal)
Langue:anglais
Publié: 15 April 2026
In: Journal of ovarian research
Year: 2026, Pages: 1-41
ISSN:1757-2215
DOI:10.1186/s13048-026-02062-5
Accès en ligne:Verlag, kostenfrei, Volltext: https://doi.org/10.1186/s13048-026-02062-5
Verlag, kostenfrei, Volltext: https://ovarianresearch.biomedcentral.com/articles/10.1186/s13048-026-02062-5
Accéder au texte intégral
Notes sur l'auteur:Saber SamadiAfshar, Hossein Azizi, Mahla Masoudi, Sahel SamadiAfshar, Ali Nikakhtar & Thomas Skutella
Description
Résumé:Ovarian cancer remains one of the most lethal gynecological malignancies, with borderline ovarian tumors (BOTs) representing lesions of low malignant potential requiring distinct surgical management. Accurate preoperative differentiation between BOTs and epithelial ovarian cancer (OC) is essential for optimizing treatment planning and patient counseling. We developed machine learning models to distinguish BOTs from OC using routine clinical data from 404 patients (199 BOTs, 205 OC) comprising 49 demographic, clinical, and biochemical parameters. Data were obtained from a public repository (n=349) and a single institutional cohort (n=55). We evaluated five algorithms: Random Forest, Support Vector Machine, Neural Network, Logistic Regression, and Decision Tree, using stratified train-test splits and stratified 5-fold cross-validation repeated five times to ensure robust performance estimation. The Random Forest model achieved the highest performance with an area under the receiver operating characteristic curve (AUC-ROC) of 0.95 (95% CI: 0.92–0.98), accuracy of 0.91 (95% CI: 0.87–0.94), sensitivity of 0.88 (95% CI: 0.83–0.92), and specificity of 0.94 (95% CI: 0.90–0.97) at the optimal threshold determined by Youden's index. Feature importance analysis identified HE4 (weight=0.224), CA125 (weight=0.089), and neutrophil count (weight=0.072) as the most discriminative predictors. Performance was comparable across both data sources, with no significant domain shift detected. Machine learning analysis of readily available laboratory parameters demonstrates potential for preoperative differentiation of BOTs from OC. A web-based prototype tool has been developed to facilitate future validation studies.
Description:Gesehen am 02.06.2026
Description matérielle:Online Resource
ISSN:1757-2215
DOI:10.1186/s13048-026-02062-5