Related Experiment Video
Updated: Jul 17, 2026

An Integrated Workflow of Identification and Quantification on FDR Control-Based Untargeted Metabolome
Published on: September 20, 2022
Federated feature selection with false discovery rate control
Jie Hu1, Jiayi Tong1,2, Yang Ning3
1Department of Biostatistics, Epidemiology & Informatics, University of Pennsylvania, Philadelphia, PA, USA.
Abstract:
Selecting a set of universally relevant features associated with a given response variable across multiple distributed data sites is an important problem in numerous scientific fields. However, performing this federated feature selection task becomes challenging when individual-level data cannot be shared due to privacy concerns. The problem is further complicated by potential heterogeneity in both feature distributions and model parameters across sites. In this paper, we propose Fed-false discovery rate (FDR), a federated feature selection framework that simultaneously identifies important features while controlling the FDR. To ensure privacy preservation and reduce communication costs, the Fed-FDR shares only lower-dimensional coefficient estimates instead of transmitting summary statistics for all features, with the dimensionality shown to be of the same order as the number of relevant features. The coordinating centre then leverages these lower-dimensional coefficient estimates to construct a generalized mirror statistic to identify the important features. The Fed-FDR is robust to the heterogeneity of feature distribution and model parameters, easy to implement, and computationally efficient. We further demonstrate that Fed-FDR effectively controls the FDR while achieving strong statistical power in our simulation studies. The results of the empirical study also demonstrate that the method is both valid and implementation-ready.
Related Concept Videos
Identifying Statistically Significant Differences: The F-Test
Frequency-dependent Selection
Fisher's Exact Test
Expected Frequencies in Goodness-of-Fit Tests
Quantifying and Rejecting Outliers: The Grubbs Test
Factorial Design
