Related Experiment Video
Updated: Jul 28, 2026

14:34
How to Create and Use Binocular Rivalry
Published on: November 10, 2010
Goodman and Kruskal's Gamma Coefficient for Ordinalized Bivariate Normal Distributions
Alessandro Barbiero1, Asmerilda Hitaj2
1Department of Economics, Management, and Quantitative Methods, Università degli Studi di Milano, Via Conservatorio, 7, 20122, Milan, Italy. alessandro.barbiero@unimi.it.
Psychometrika
|October 27, 2020
Summary
Discretizing a bivariate normal distribution affects ordinal association measures. Goodman and Kruskal
Area of Science:
- Statistics
- Probability Theory
- Ordinal Data Analysis
Background:
- Bivariate normal distributions are fundamental in statistical modeling.
- Ordinal association measures like Goodman and Kruskal's gamma are crucial for analyzing ranked data.
- Discretization transforms continuous variables into ordinal categories, potentially altering association metrics.
Purpose of the Study:
- To investigate the impact of discretization on ordinal association measures derived from a bivariate normal distribution.
- To compare Goodman and Kruskal's gamma with Kendall's rank correlation before and after discretization.
- To propose a method for constructing bivariate ordinal variables with specific marginal distributions and association levels.
Main Methods:
- Considered a bivariate normal distribution with linear correlation.
- Discretized the continuous components using assigned sets of thresholds.
- Calculated Goodman and Kruskal's gamma and Kendall's rank correlation for the resulting ordinal variables.
- Explored various experimental settings by varying thresholds and marginal distributions.
Main Results:
- The absolute value of Goodman and Kruskal's gamma was consistently higher than Kendall's rank correlation after discretization.
- This discrepancy decreased with an increased number of categories.
- Equally probable categories also reduced the difference between the two association measures.
Conclusions:
- Discretization systematically inflates the measured ordinal association compared to the continuous case.
- The extent of this inflation depends on the number of categories and the distribution of thresholds.
- A method is proposed for creating bivariate ordinal variables with controlled marginals and association by ordinalizing a bivariate normal distribution.
Related Concept Videos
Normal Distribution
The normal, a continuous distribution, is the most important of all the distributions. Its graph is a bell-shaped symmetrical curve, which is observed in almost all disciplines. Some of these include psychology, business, economics, the sciences, nursing, and, of course, mathematics. Some instructors may use the normal distribution to help determine students’ grades. Most IQ scores are normally distributed. Often real-estate prices fit a normal distribution. The normal distribution is extremely...
Goodness-of-Fit Test
The goodness-of-fit test is a type of hypothesis test which determines whether the data "fits" a particular distribution. For example, one may suspect that some anonymous data may fit a binomial distribution. A chi-square test (meaning the distribution for the hypothesis test is chi-square) can be used to determine if there is a fit. The null and alternative hypotheses may be written in sentences or stated as equations or inequalities. The test statistic for a goodness-of-fit test is given as...
Expected Frequencies in Goodness-of-Fit Tests
A goodness-of-fit test is conducted to determine whether the observed frequency values are statistically similar to the frequencies expected for the dataset. Suppose the expected frequencies for a dataset are equal such as when predicting the frequency of any number appearing when casting a die. In that case, the expected frequency is the ratio of the total number of observations (n) to the number of categories (k).
Wald-Wolfowitz Runs Test II
The Wald-Wolfowitz runs test, commonly referred to as the runs test, is a nonparametric test used to assess the randomness of ordered data. The test evaluates the number of runs, which are consecutive sequences of similar elements within the data. If the number of runs is significantly higher or lower than expected, the data is considered non-random, indicating a detectable pattern or structure.
For binary data, runs are identified using symbols such as + and −, or equivalently, 1s and 0s. In...
For binary data, runs are identified using symbols such as + and −, or equivalently, 1s and 0s. In...
One-Compartment Open Model: Wagner-Nelson and Loo Riegelman Method for ka Estimation
This lesson introduces two critical methods in pharmacokinetics, the Wagner-Nelson and Loo-Riegelman methods, used for estimating the absorption rate constant (ka) for drugs administered via non-intravenous routes. The Wagner-Nelson method relates ka to the plasma concentration derived from the slope of a semilog percent unabsorbed time plot. However, it is limited to drugs with one-compartment kinetics and can be impacted by factors like gastrointestinal motility or enzymatic degradation.
On...
On...
Friedman Two-way Analysis of Variance by Ranks
Friedman's Two-Way Analysis of Variance by Ranks is a nonparametric test designed to identify differences across multiple test attempts when traditional assumptions of normality and equal variances do not apply. Unlike conventional ANOVA, which requires normally distributed data with equal variances, Friedman's test is ideal for ordinal or non-normally distributed data, making it particularly useful for analyzing dependent samples, such as matched subjects over time or repeated measures from...

