Related Experiment Video
Updated: May 10, 2025

Comparing the Frequency Effect Between the Lexical Decision and Naming Tasks in Chinese
Published on: April 1, 2016
A Study of the Spatial Correlogram Patterns of Chinese Surnames
Xiaohui Fan1, Xuemin Zhang2, Yuan Gao1
1School of Systems Science, Beijing Normal University, Beijing, China.
Objectives:
This study investigates the historical diffusion and migration patterns of Chinese surnames by analyzing their spatial correlograms. The primary objectives are to identify typical correlogram categories, characterize each category, and explore the factors influencing the historical diffusion and migration processes that have shaped the spatial distributions of Chinese surnames.
Data And Methods:
The data used in this study come from China's National Citizen Identity Information Center (NCIIC), which provides surname and prefecture information for 1.28 billion individuals. We calculate spatial correlograms to assess surname autocorrelation across varying geographic distances and apply cluster analysis to classify the 380 most common surnames, covering 97% of the population, into five categories based on their spatial correlograms. We examine the characteristics of correlograms across these categories and propose an index to capture the overall geographic distribution of surnames in a category.
Results:
In the analysis, five distinct categories of spatial correlograms are identified: C (cline), SC (slight cline), IBD (isolation by distance), D (depression), and IBD + D (isolation by distance + depression). Surnames in category C exhibit a broad and even distribution, with high autocorrelation in adjacent regions and a large diffusion range. Surnames in category SC show lower autocorrelation than those in category C but still exhibit a large diffusion range. Surnames in category IBD are highly concentrated in specific regions, with low autocorrelation and a smaller diffusion range. Surnames in both categories D and IBD + D display long-distance autocorrelation, featuring a distinct depression in their correlograms.
Discussion:
Surnames with long histories and significant influence, such as those in category C, tend to be broadly and evenly distributed, reflecting prolonged diffusion processes. Conversely, surnames with more recent origins or those that have experienced isolation, such as those in category IBD, typically exhibit more concentrated distributions. The study also highlights the role of large-scale, long-distance migration events in shaping Chinese surname distributions, particularly for surnames in categories D and IBD + D.
Related Concept Videos
Residual Plots
When the residual values are plotted against the variable x, it is called a residual...
Correlation
Two variables, for example, a and b, are said to be positively correlated if both variables move in the same direction. In other words, a positive correlation exists between two variables, a and b, if:
Spearman's Rank Correlation Test
Spearman's test calculates...
Scatter Plot
Correlations
2D NMR: Overview of Homonuclear Correlation Techniques
COSY90 is the standard two-dimensional (2D) COSY experiment that...

