Related Experiment Video
Updated: Jun 30, 2026

Selecting Multiple Biomarker Subsets with Similarly Effective Binary Classification Performances
Published on: October 11, 2018
Speeding Up the Discovery of Optimal Feature Combinations for Omics Data Based on Pseudo-Kernel Function
Shipeng Ren1,2, Guoqing Yang1,2, Deyin Yu1,2
1Dalian Key Laboratory of Smart Fisheries, Dalian 116023, Liaoning Province, P. R. China.
Abstract:
Discovering meaningful feature (molecule) combinations to define simple, accurate, and easily interpretable decision rules for disease classification and prediction can improve the study of disease diagnosis and prognosis. However, the computational time complexity of constructing feature combinations for each feature pair in existing methods is often problematic or prohibitive, as the number of features is often in the order of tens of thousands. To significantly reduce the computational cost and maintain the classification performance, this paper proposed a novel acceleration algorithm and a new omics data analysis method based on pseudo kernel functions (PKF- -TSP). PKF- -TSP explores the linear and nonlinear combination of features by pseudo kernel function, evaluate feature interaction, and selects top-scoring pairs to build an ensemble classifier. PKF- -TSP maps feature pairs from a low-dimensional space to a high-dimensional feature space by a mapping function, as the same effect as kernel function, and ensure the classification performance. However, it significantly reduces the computational time costs. Experimental results demonstrate that PKF- -TSP achieves superior classification performance, while exhibiting significantly improved computational efficiency compared with KF- -TSP, with the running time reduced by 72.43%. Furthermore, the feature pairs identified by PKF- -TSP align with physiological and pathological changes, offering insights into disease mechanisms. The method also excels in cross-cancer pathway interaction analysis, capturing both conserved and tissue-specific signaling networks. Hence, PKF- -TSP enables rapid and efficient feature mining, which is especially suitable for large-scale disease omics data analysis.
Related Concept Videos
Methods of Medium Optimization
Model-Independent Approaches for Pharmacokinetic Data: Noncompartmental Analysis
One important characteristic of noncompartmental analyses is that drug exposure increases proportionally with increasing doses. This relationship...
Pharmacokinetic Models: Comparison and Selection Criterion
Physiological models take a detailed approach by considering specific molecular processes. They can predict drug distribution, metabolism, and elimination changes, providing a comprehensive understanding of how drugs interact with the body.
Model Approaches for Pharmacokinetic Data: Distributed Parameter Models
The distributed parameter models are specifically designed to account for variations and differences in some drug classes. This model is particularly useful for assessing regional concentrations of anticancer or...
One-Compartment Open Model: Wagner-Nelson and Loo Riegelman Method for ka Estimation
On...
Optimization Problems