Related Experiment Video
Updated: May 23, 2025

Selecting Multiple Biomarker Subsets with Similarly Effective Binary Classification Performances
Published on: October 11, 2018
A stacked learning framework for accurate classification of polycystic ovary syndrome with advanced data balancing
Heba M Emara1, Walid El-Shafai2, Naglaa F Soliman3
1Department of Electronics and Electrical Communications Engineering, Ministry of Higher Education Pyramids Higher Institute (PHI) for Engineering and Technology, 6th of October City, Egypt.
Introduction:
In the domain of women's health, the intricate conditions of Polycystic Ovary Syndrome (PCOS) demand sophisticated methodologies for accurate identification and intervention.
Methods:
This study introduces an innovative machine learning framework tailored to precisely classify instances of PCOS. The methodology incorporates stacked learning and depends on the Adaptive Synthetic (ADASYN) algorithm, Synthetic Minority Over-sampling Technique (SMOTE), and random oversampling methods for addressing data imbalances. The BORUTA technique is used for feature selection, with the overarching objective of advancing precision and performance metrics in classification tasks.
Results:
Within the scope of PCOS classification, the proposed framework achieves a commendable 97% accuracy. These results underscore the proficiency of the proposed framework in discriminating PCOS cases with a high degree of precision. Critical to this contribution is the rigorous comparative analysis against existing methodologies, affirming the superior accuracy and performance attributes of the proposed framework.
Discussion:
This substantiates its potential as a transformative tool in medical classification. Moreover, beyond immediate applications, this paper explores the generalization of the proposed framework, demonstrating its adaptability and efficacy across different medical classifications. This versatility is exemplified by its successful application to cervical cancer, showcasing the framework potential as a pioneering force in reshaping the landscape of machine-learning applications in healthcare diagnostics.
Related Concept Videos
Classification of Systems-II
Classification of Systems-I
Homogeneity dictates that if an input x(t) is multiplied by a constant c, the output y(t) is multiplied by the same constant. Mathematically, this is expressed as:
Statistical Software for Data Analysis and Clinical Trials
Aggregates Classification
Petrographic classification groups aggregates based on common mineralogical characteristics. Some of the common mineral groups found in aggregates are...
Multiple Regression
Farmers can use multiple regression to determine the crop yield based on more than one factor, such as water availability, fertilizer, soil properties, etc. Here, the crop yield is the response or dependent variable as it depends on the other independent variables. The analysis requires the construction of a scatter plot...
Classification of Leukocytes
Neutrophils are the most abundant type of granular leukocytes, comprising 50-70% of all leukocytes. They feature small, evenly distributed granules and a...

