Related Experiment Video
Updated: Sep 10, 2025

Selecting Multiple Biomarker Subsets with Similarly Effective Binary Classification Performances
Published on: October 11, 2018
Improving learning from the complex multi-class imbalanced and overlapped data by mapping into higher dimension using
Zafar Mahmood1, Leila Jamel2, Dina Ahmed Salem3
1Department of Computer Science, University of Gujrat, Gujrat, Pakistan.
None:
Several issues are there to prevent the traditional classifiers from getting an acceptable performance level while learning from multi-class problems. One of the main problems is the unequal distribution of samples, which significantly reduces the efficiency of the underlying classifier when combined with incompatible optimization benchmarks and data overlapping phenomena. The classifier performance is compromised beyond the expected level by the combined effects of imbalanced distribution and sample overlapping around the class boundaries. This problem worsens with the increase in the number of classes in the multi-class scenario. Despite having a more significant combined effect on classifier performance, the combined effects of imbalanced data and overlapping questions have been given the least attention in the research. To improve models' learning from imbalanced multi-class and overlapping of shared attributes issues, this work introduces SVM++, a modified version of support vector machines (SVM). Comprising of three steps, Algorithm-1 finds and splits the training set into overlapping and non-overlapping samples. Algorithm-2 then separates the overlapped data into the Critical-1 and Critical-2 regions. The Critical-1 region consists of overlapped samples, sharing similar characteristics, which is the main cause of degraded classification performance. In the third step, an algorithm based on the mean of the maximum and minimum distance of the Critical-1 region samples is proposed by improving the traditional SVM kernel mapping function to a higher dimension. Thirty real datasets with various imbalances and degrees of overlap are utilized to compare our suggested algorithms' supremacy with the state-of-the-art classifiers.
More Related Videos
Related Concept Videos
Classification of Systems-II
Classification of Systems-I
Homogeneity dictates that if an input x(t) is multiplied by a constant c, the output y(t) is multiplied by the same constant. Mathematically, this is expressed as:
Aggregates Classification
Petrographic classification groups aggregates based on common mineralogical characteristics. Some of the common mineral groups found in aggregates are...
Multi-input and Multi-variable systems
In the absence...
Multiple Regression
Farmers can use multiple regression to determine the crop yield based on more than one factor, such as water availability, fertilizer, soil properties, etc. Here, the crop yield is the response or dependent variable as it depends on the other independent variables. The analysis requires the construction of a scatter plot...
Associative Learning
Classical conditioning, also known...

