Ethnicity-based name partitioning for author name disambiguation using supervised machine learning.

Jinseok Kim1, Jenna Kim2, Jason Owen-Smith3

  • 1Institute for Research on Innovation & Science, Survey Research Center, Institute for Social Research University of Michigan Ann Arbor Michigan USA.

Summary

Partitioning author names by ethnicity significantly improves disambiguation accuracy. Tailoring machine learning models to specific ethnic name groups enhances performance, outperforming general approaches.

Related Concept Videos

Pedigree Analysis01:35

Pedigree Analysis

Overview
86.2K
Ethnic Identity within a Larger Culture01:27

Ethnic Identity within a Larger Culture

Adolescents from ethnic minority backgrounds face a multifaceted journey in forming their identities, shaped by the intersections of cultural expectations and personal exploration. For these adolescents, identity formation involves not only typical developmental challenges but also navigating the perceptions and attitudes of the majority culture. As they grow, adolescents in ethnic minority groups often become increasingly aware of stereotypes, social biases, and discrimination, all of which...
121
Law of Independent Assortment02:03

Law of Independent Assortment

While Mendel’s Law of Segregation states that the two alleles for one gene are separated into different gametes, a different question of how different genes are inherited remains. For example, is the gene for tall plants inherited with the gene for green peas? Mendel asked this question by experimenting with a dihybrid cross; a cross in which both parents are homozygous for two distinct traits resulting in an F1 generation that are heterozygous for both traits.
59.4K
Multiple Allele Traits01:49

Multiple Allele Traits

The Concept of Multiple Allelism
36.1K
Genetic Lingo01:11

Genetic Lingo

Overview
108.0K