Vincent Schilling, Peter Beyerlein, Jeremy Chien

A Bioinformatics Analysis of Ovarian Cancer Data Using Machine Learning

  • Computational Mathematics
  • Computational Theory and Mathematics
  • Numerical Analysis
  • Theoretical Computer Science

The identification of biomarkers is crucial for cancer diagnosis, understanding the underlying biological mechanisms, and developing targeted therapies. In this study, we propose a machine learning approach to predict ovarian cancer patients’ outcomes and platinum resistance status using publicly available gene expression data. Six classical machine-learning algorithms are compared on their predictive performance. Those with the highest score are analyzed by their feature importance using the SHAP algorithm. We were able to select multiple genes that correlated with the outcome and platinum resistance status of the patients and validated those using Kaplan–Meier plots. In comparison to similar approaches, the performance of the models was higher, and different genes using feature importance analysis were identified. The most promising identified genes that could be used as biomarkers are TMEFF2, ACSM3, SLC4A1, and ALDH4A1.

Need a simple solution for managing your BibTeX entries? Explore CiteDrive!

  • Web-based, modern reference management
  • Collaborate and share with fellow researchers
  • Integration with Overleaf
  • Comprehensive BibTeX/BibLaTeX support
  • Save articles and websites directly from your browser
  • Search for new articles from a database of tens of millions of references
Try out CiteDrive

More from our Archive