维基百科的文章类别如何决定其编辑者的异质性
Aileen Oeberst1,2, Till Ridderbecks3
1Department of Psychology, University of Hagen, 58084, Hagen, Germany. aileen.oeberst@fernuni-hagen.de.
Scientific reports
|January 7, 2024
概括
维基百科编辑倾向于自行选择文章. 专门与特定国家相关的主题吸引了该国更高比例的编辑,这可能会影响文章的多样性和质量.
科学领域:
- 社会科学 社会科学 社会科学
- 信息科学 信息科学 信息科学
- 计算机介导通信 计算机介导通信
背景情况:
- 合作对于知识进步和社会进步至关重要.
- 网络2.0技术使前所未有的协作创作成为可能,维基百科就是一个例子.
- 维基百科的文章质量与编辑数量和多样性相关.
研究的目的:
- 调查维基百科编辑中自我选择偏见的潜在威胁.
- 检查是否国家归属影响编辑对特定国家文章的选择.
主要方法:
- 对维基百科多种语言版本的编辑人口统计和文章内容的分析.
- 对文章的编辑国籍比例进行比较,这些文章具有独家的国家链接,而不是具有通用或国际链接的文章.
- 基于文章时事性的编辑自选模式的统计检查.
主要成果:
- 在维基百科中观察到自我选择的显著效应.
- 与某个特定国家的专属链接的文章吸引了来自该国家的编辑比例不成比例地高.
- 来自特定国家的编辑的比例随着文章的国家链接的独家性而增加.
结论:
- 基于国家身份的维基百科编辑的自我选择是一个有记录的现象.
- 这种国家自选可能会威胁到编辑的多样性,从而影响文章的质量.
- 需要进一步的研究来理解和减轻编辑自选对协作知识平台的影响.
相关概念视频
What is Biodiversity?
Biodiversity describes the variety of living things at multiple organizational levels: genetic, species and ecosystem diversity. Species diversity includes all branches of the evolutionary tree from single-celled prokaryotic organisms, bacteria, and archaea, to the eukaryotic kingdoms: plants; animals; fungi; and protists. To date, there have been about 1.75 million species identified, and new species are discovered every week.
Types of Selection
Natural selection influences the frequencies of particular alleles and phenotypes within populations in several different ways. Primarily, natural selection can be directional, stabilizing, or disruptive. Directional selection favors one extreme trait and shifts the population towards that phenotype while selecting against individuals displaying alternate traits. Stabilizing selection favors an intermediate trait with a narrow range of variation. Deviation from the optimal phenotype towards an...
How Data are Classified: Categorical Data
A variable, usually notated by capital letters such as X and Y, is a characteristic or measurement that can be determined for each member of a population. Data are the actual values of variables. They may be numbers, or they may be words. Datum is a single value.
Data are classified based on whether they are measurable or not. Categorical data cannot be measured; instead, it can be divided into categories. For example, if Y denotes a person's party affiliation, some examples of Y include...
Data are classified based on whether they are measurable or not. Categorical data cannot be measured; instead, it can be divided into categories. For example, if Y denotes a person's party affiliation, some examples of Y include...
Test for Homogeneity
The goodness–of–fit test can be used to decide whether a population fits a given distribution, but it will not suffice to decide whether two populations follow the same unknown distribution. A different test, called the test for homogeneity, can be used to conclude whether two populations have the same distribution. To calculate the test statistic for a test for homogeneity, follow the same procedure as with the test of independence. The hypotheses for the test for homogeneity can be stated as...
Categories of Equilibrium
Equilibrium is a crucial concept in physics, enabling us to understand how forces interact with bodies to produce no or constant motion. In two-dimensional equilibrium, force systems can be classified into different categories based on their characteristics.
One of the categories of equilibrium is collinear equilibrium, which involves forces acting along a straight line. This type of equilibrium requires only one force equation in the direction of the forces, as the other equations are...
One of the categories of equilibrium is collinear equilibrium, which involves forces acting along a straight line. This type of equilibrium requires only one force equation in the direction of the forces, as the other equations are...
Variability: Analysis
Measures of variability are statistical metrics that reveal the dispersion pattern within a dataset. They are pivotal in biostatistics, providing insights into the heterogeneity within health and biological data. Variability signifies the degree to which data points diverge from one another, helping researchers understand the potential range of values and associated uncertainty within the data.
The range is a simple measure of variability, indicating the difference between the highest and...
The range is a simple measure of variability, indicating the difference between the highest and...


