Related Experiment Video
Updated: Apr 28, 2026

A Method for Quantifying Foliage-Dwelling Arthropods
Published on: October 20, 2019
An improved nonparametric lower bound of species richness via a modified good-turing frequency formula
Chun-Huo Chiu1, Yi-Ting Wang1, Bruno A Walther2
1Institute of Statistics, National Tsing Hua University, Hsin-Chu 30043, Taiwan.
Abstract:
It is difficult to accurately estimate species richness if there are many almost undetectable species in a hyper-diverse community. Practically, an accurate lower bound for species richness is preferable to an inaccurate point estimator. The traditional nonparametric lower bound developed by Chao (1984, Scandinavian Journal of Statistics 11, 265-270) for individual-based abundance data uses only the information on the rarest species (the numbers of singletons and doubletons) to estimate the number of undetected species in samples. Applying a modified Good-Turing frequency formula, we derive an approximate formula for the first-order bias of this traditional lower bound. The approximate bias is estimated by using additional information (namely, the numbers of tripletons and quadrupletons). This approximate bias can be corrected, and an improved lower bound is thus obtained. The proposed lower bound is nonparametric in the sense that it is universally valid for any species abundance distribution. A similar type of improved lower bound can be derived for incidence data. We test our proposed lower bounds on simulated data sets generated from various species abundance models. Simulation results show that the proposed lower bounds always reduce bias over the traditional lower bounds and improve accuracy (as measured by mean squared error) when the heterogeneity of species abundances is relatively high. We also apply the proposed new lower bounds to real data for illustration and for comparisons with previously developed estimators.
More Related Videos
Related Concept Videos
Expected Frequencies in Goodness-of-Fit Tests
Goodness-of-Fit Test
Relative Frequency Distribution
Frequency-dependent Selection
Construction of Frequency Distribution
First, make a table with two columns—one with the title of the data that needs to be organized, and the other column for frequency. [Draw a third column for tally marks if needed]. Then, take a look at the items given in the data set and decide if an ungrouped frequency distribution table or a grouped frequency distribution table would be more suitable. If there are large sets of different values, then it is...
Quantifying and Rejecting Outliers: The Grubbs Test

