Related Experiment Video
Updated: Mar 22, 2026

05:02
Comparing Bibliometric Analysis Using PubMed, Scopus, and Web of Science Databases
Published on: October 24, 2019
34.3K
Clustering Scientific Publications Based on Citation Relations: A Systematic Comparison of Different Methods
Lovro Šubelj1, Nees Jan van Eck2, Ludo Waltman2
1University of Ljubljana, Faculty of Computer and Information Science, Ljubljana, Slovenia.
Plos One
|April 29, 2016
Summary
This study compares bibliometric clustering methods for identifying research fields. Map equation methods showed the best performance in clustering publications within citation networks.
Area of Science:
- Bibliometrics
- Network Science
- Data Science
Background:
- Clustering methods are crucial for identifying research areas in bibliometric studies.
- These methods group publications based on citation network relationships.
- Numerous graph partitioning and community detection techniques exist in network science.
Purpose of the Study:
- To systematically compare the performance of various clustering methods for citation networks.
- To analyze the statistical properties and expert-based assessment of clustering results.
- To identify optimal methods for research field identification in bibliometrics.
Main Methods:
- Systematic comparison of numerous clustering algorithms on diverse citation networks (small to large scale).
- Statistical analysis of clustering results.
- Expert-based evaluation focusing on scientometrics publications.
Main Results:
- A trade-off exists between desirable properties for effective publication clustering.
- Map equation methods demonstrated superior performance in the comparative analysis.
- Findings suggest map equation methods warrant increased attention from bibliometric researchers.
Conclusions:
- Map equation methods offer a promising approach for research field identification.
- The study provides valuable insights into the strengths and weaknesses of different bibliometric clustering techniques.
- Further investigation into map equation methods is recommended for the bibliometric community.
More Related Videos
Related Concept Videos
Cluster Sampling Method
15.4K
Appropriate sampling methods ensure that samples are drawn without bias and accurately represent the population. Because measuring the entire population in a study is not practical, researchers use samples to represent the population of interest.
To choose a cluster sample, divide the population into clusters (groups) and then randomly select some of the clusters. All the members from these clusters are in the cluster sample. For example, if you randomly sample four departments from your...
To choose a cluster sample, divide the population into clusters (groups) and then randomly select some of the clusters. All the members from these clusters are in the cluster sample. For example, if you randomly sample four departments from your...
15.4K
Protein Networks
2.9K
2.9K
Protein Networks
4.7K
An organism can have thousands of different proteins, and these proteins must cooperate to ensure the health of an organism. Proteins bind to other proteins and form complexes to carry out their functions. Many proteins interact with multiple other proteins creating a complex network of protein interactions.
These interactions can be represented through maps depicting protein-protein interaction networks, represented as nodes and edges. Nodes are circles that are representative of a protein,...
These interactions can be represented through maps depicting protein-protein interaction networks, represented as nodes and edges. Nodes are circles that are representative of a protein,...
4.7K
Chi-square Analysis
44.7K
The chi-square test is a statistical hypothesis test. It is used to check whether there is a significant difference between an expected value and an observed value. In the context of genetics, it enables us to either accept or reject a hypothesis, based on how much the observed values deviate from the expected values.
The chi-square test was developed by Pearson in 1990.
The first step of performing a Chi-square analysis is to establish a null hypothesis, which assumes that there is no real...
The chi-square test was developed by Pearson in 1990.
The first step of performing a Chi-square analysis is to establish a null hypothesis, which assumes that there is no real...
44.7K
Network Covalent Solids
16.5K
Network covalent solids contain a three-dimensional network of covalently bonded atoms as found in the crystal structures of nonmetals like diamond, graphite, silicon, and some covalent compounds, such as silicon dioxide (sand) and silicon carbide (carborundum, the abrasive on sandpaper). Many minerals have networks of covalent bonds.
To break or to melt a covalent network solid, covalent bonds must be broken. Because covalent bonds are relatively strong, covalent network solids are typically...
To break or to melt a covalent network solid, covalent bonds must be broken. Because covalent bonds are relatively strong, covalent network solids are typically...
16.5K
Outliers and Influential Points
6.6K
An outlier is an observation of data that does not fit the rest of the data. It is sometimes called an extreme value. When you graph an outlier, it will appear not to fit the pattern of the graph. Some outliers are due to mistakes (for example, writing down 50 instead of 500), while others may indicate that something unusual is happening. Outliers are present far from the least squares line in the vertical direction. They have large "errors," where the "error" or residual is the...
6.6K

