Related Experiment Video
Updated: May 15, 2026

An Integrated Workflow of Identification and Quantification on FDR Control-Based Untargeted Metabolome
Published on: September 20, 2022
When target-decoy false discovery rate estimations are inaccurate and how to spot instances
1Department of Pharmaceutical Chemistry, University of California San Francisco , 600 16th Street, Genentech Hall Room N474A, San Francisco, California 94158, USA. chalkley@cgl.ucsf.edu
Abstract:
To address problems with estimating the reliability of proteomic search engine results from mass spectrometry fragmentation data, the use of target-decoy database searching has become the de facto approach for estimating a false discovery rate. Several articles have been written about the effects of different ways of creating the decoy database, effects of the search engine scoring, or effects of search parameters on whether this approach provides an accurate estimate, not all agreeing with each other's conclusions. Hence, there may be some confusion about how effective this approach is and how broadly it can be applied. Although it is generally very effective, in this article I will try to emphasize some of the pitfalls and dangers of using the target-decoy approach and will indicate tell-tale signs that something may be amiss. This information will hopefully help researchers become more astute in their assessment of search results.
Related Concept Videos
Understanding Deception
False Memories
One primary source of false memories is misattribution, where individuals incorrectly associate external information with...
Accuracy and Errors in Hypothesis Testing
In hypothesis testing, the probability of making a Type I error, denoted as α, is commonly set at 0.05. This significance level indicates a 5% chance...
Detection of Gross Error: The Q Test
Difference from Background: Limit of Detection
The LOD indicates the presence or absence...
Testing a Claim about Mean: Unknown Population SD
Estimating a population mean requires the samples to be approximately normally distributed. The data should be collected from the randomly selected samples having no sampling bias. There is no specific requirement for sample size. But if the sample size is less than 30, and we don't know the population standard deviation, a different approach is used; instead...
