Search research articles

关于 JoVE

概览领导团队博客 JoVE 帮助中心

作者

出版流程编辑委员会范围与政策同行评审常见问题投稿

图书馆员

用户评价订阅访问资源图书馆顾问委员会常见问题

研究

JoVE Journal Methods Collections JoVE Encyclopedia of Experiments 存档

教育

JoVE Core JoVE Business JoVE Science Education JoVE Lab Manual 教师资源中心教师网站

使用条款与条件

相关概念视频

Quantifying and Rejecting Outliers: The Grubbs Test

Quantifying and Rejecting Outliers: The Grubbs Test

Sometimes, a data set can have a recorded numerical observation that greatly deviates from the rest of the data. Assuming that the data is normally distributed, a statistical method called the Grubbs test can be used to determine whether the observation is truly an outlier. To perform a two-tailed Grubbs test, first, calculate the absolute difference between the outlier and the mean. Then, calculate the ratio between this difference and the standard deviation of the sample. This...

Weighted Mean

Weighted Mean

While taking the arithmetic, geometric, or harmonic mean of a sample data set, equal importance is assigned to all the data points. However, all the values may not always be equally important in some data sets. An intrinsic bias might make it more important to give more weightage to specific values over others.
For example, consider the number of goals scored in the matches of a tournament. While computing the average number of goals scored in the tournament, it may be more important to...

Expected Frequencies in Goodness-of-Fit Tests

Expected Frequencies in Goodness-of-Fit Tests

A goodness-of-fit test is conducted to determine whether the observed frequency values are statistically similar to the frequencies expected for the dataset. Suppose the expected frequencies for a dataset are equal such as when predicting the frequency of any number appearing when casting a die. In that case, the expected frequency is the ratio of the total number of observations (n) to the number of categories (k).

Classification of Signals

Classification of Signals

In signal processing, signals are classified based on various characteristics: continuous-time versus discrete-time, periodic versus aperiodic, analog versus digital, and causal versus noncausal. Each category highlights distinct properties crucial for understanding and manipulating signals.
A continuous-time signal holds a value at every instant in time, representing information seamlessly. In contrast, a discrete-time signal holds values only at specific moments, often denoted as x(n), where...

Sensitivity, Specificity, and Predicted Value

Sensitivity, Specificity, and Predicted Value

In healthcare diagnostics, laboratory tests play a crucial role in identifying and diagnosing a wide range of medical conditions. However, interpreting test results is not always straightforward. An abnormal test result does not always confirm the presence of a disease, just as a normal result does not guarantee its absence. To assess the reliability of these diagnostic tools, healthcare practitioners rely on two key statistical indicators: sensitivity and specificity.
Sensitivity is the...

Frequency-dependent Selection

Frequency-dependent Selection

When the fitness of a trait is influenced by how common it is (i.e., its frequency) relative to different traits within a population, this is referred to as frequency-dependent selection. Frequency-dependent selection may occur between species or within a single species. This type of selection can either be positive—with more common phenotypes having higher fitness—or negative, with rarer phenotypes conferring increased fitness.

您也可能阅读

相关文章

通过共同作者、期刊和引用图与本文相关的文章。

排序

Same author

Employing an immunoinformatics approach revealed potent multi-epitope based subunit vaccine for lymphocytic choriomeningitis virus.

Journal of infection and public health·2023

Same author

Microbial-induced calcium carbonate precipitation: Influencing factors, nucleation pathways, and application in waste water remediation.

The Science of the total environment·2022

Same author

Synergistic removal of nitrate by a cellulose-degrading and denitrifying strain through iron loaded corn cobs filled biofilm reactor at low C/N ratio: Capability, enhancement and microbiome analysis.

Bioresource technology·2022

Same author

Ring Expansion of Isatins <i>via</i> 1,2-Phospha-Brook Rearrangement: A Route to the Synthesis of 2-Quinolinone-Derived <i>p</i>-Quinone Methides.

The Journal of organic chemistry·2022

Same author

Production of a recyclable nanobiocatalyst to synthesize quinazolinone derivatives.

RSC advances·2022

Same author

Comparative Pan-Genomic Analysis Revealed an Improved Multi-Locus Sequence Typing Scheme for <i>Staphylococcus aureus</i>.

Genes·2022

Same journal

Novel Parent Survey Measures Sensory Behaviors Incorporating Sensory Modality and Stimulus Intensity.

Heliyon·2026

Same journal

Expression of concern: "SQSTM1/p62 promotes the progression of gastric cancer through epithelial-mesenchymal transition" [Heliyon 10 (2024) e24409].

Heliyon·2026

Same journal

Expression of concern: "TL1A promotes metastasis and EMT process of colorectal cancer" [Heliyon 10 (2024) e24392].

Heliyon·2026

Same journal

Expression of concern: "Factors affecting timing of surgery following neoadjuvant chemoradiation for esophageal cancer" [Heliyon 9 (2023) e23212].

Heliyon·2026

Same journal

Expression of concern: "On stratified single-valued soft topogenous structures" [Heliyon 10 (2024) e27926].

Heliyon·2026

Same journal

Expression of concern: "Artifact removal and motor imagery classification in EEG using advanced algorithms and modified DNN" [Heliyon 10 (2024) e27198].

Heliyon·2026

查看所有相关文章

Search research articles

相关实验视频

Updated: Jun 10, 2025

Selecting Multiple Biomarker Subsets with Similarly Effective Binary Classification Performances

Selecting Multiple Biomarker Subsets with Similarly Effective Binary Classification Performances

Published on: October 11, 2018

通过高维二进制类不平衡基因表达数据的强有力的加权得分来选择特征.

Zardad Khan¹, Amjad Ali¹, Saeed Aldahmani¹

¹Department of Statistics and Business Analytics, United Arab Emirates University, Al Ain, United Arab Emirates.

|October 14, 2024

概括

此摘要是机器生成的。

一种新的特征选择方法,不平衡数据的强有力的加权得分 (ROWSU),有效地识别了不平衡基因表达数据中的关键基因. ROWSU通过选择歧视性特征来提高分类准确性,即使在偏分布中也是如此.

关键词:

功能选择选择功能选择基因表达数据基因表达数据强大的得分强大的得分.支持向量是指支持的向量.不平衡的阶级分布.

更多相关视频

Author Spotlight: Impact of Intergenic Interactions on Disease-Identifying Dark Biomarkers

Author Spotlight: Impact of Intergenic Interactions on Disease-Identifying Dark Biomarkers

Published on: March 1, 2024

Assisted Selection of Biomarkers by Linear Discriminant Analysis Effect Size LEfSe in Microbiome Data

Assisted Selection of Biomarkers by Linear Discriminant Analysis Effect Size LEfSe in Microbiome Data

Published on: May 16, 2022

相关实验视频

Last Updated: Jun 10, 2025

Selecting Multiple Biomarker Subsets with Similarly Effective Binary Classification Performances

Selecting Multiple Biomarker Subsets with Similarly Effective Binary Classification Performances

Published on: October 11, 2018

Author Spotlight: Impact of Intergenic Interactions on Disease-Identifying Dark Biomarkers

Author Spotlight: Impact of Intergenic Interactions on Disease-Identifying Dark Biomarkers

Published on: March 1, 2024

Assisted Selection of Biomarkers by Linear Discriminant Analysis Effect Size LEfSe in Microbiome Data

Assisted Selection of Biomarkers by Linear Discriminant Analysis Effect Size LEfSe in Microbiome Data

Published on: May 16, 2022

科学领域:

生物信息学是一种生物信息学.
计算生物学计算生物学
机器学习在基因组学中的应用

背景情况:

高维基因表达数据往往会带来阶级不平衡的挑战.
偏斜的类分布会对分类算法的性能产生负面影响.
有效的特征选择对于基因组学中准确的二进制分类至关重要.

研究的目的:

为特征选择提出不平衡数据 (ROWSU) 的强有力的加权得分.
为了应对高维基基因表达数据集中的阶级不平衡的挑战.
提高对不平衡的基因组数据的分类算法的性能.

主要方法:

通过合成少数人过量采样技术进行数据平衡.
对于初始最小基因子集选择的贪搜索方法.
使用支持矢量权重进行基因改进的新型加权稳健得分.

主要成果:

ROWSU方法成功地从不平衡的数据集中选择了歧视性基因.
在7个基因表达数据集上进行评估,ROWSU表现出卓越的性能.
在使用kNN和RF分类器的分类准确度,灵敏度和F1得分方面超越了现有的方法.

结论:

拟议的ROWSU方法在不平衡的基因表达数据中的特征选择中是有效的.
通过选择最具歧视性的基因,ROWSU提高了分类器的性能.
这种方法为基因组学中的二进制分类问题提供了强大的解决方案,其中包含了偏斜数据.