Related Experiment Video
Updated: Aug 30, 2025

05:02
Comparing Bibliometric Analysis Using PubMed, Scopus, and Web of Science Databases
Published on: October 24, 2019
31.9K
Research on Literature Clustering Algorithm for Massive Scientific and Technical Literature Query Service
1Wuhan University of Science & Technology Library, Wuhan 430081, Hubei, China.
Computational Intelligence and Neuroscience
|September 1, 2022
Summary
This study introduces a new R-tree indexing method for faster scientific literature retrieval in big data environments. The approach enhances both search efficiency and precision for large datasets.
Area of Science:
- Information Science
- Computer Science
- Data Science
Background:
- Traditional scientific and technical literature search methods are challenged by the exponential growth of digital information resources.
- Existing approaches struggle to meet user information needs within large-scale datasets due to limitations in traditional technical methods and service models.
Purpose of the Study:
- To propose an effective model for large-scale scientific and technical literature search services in the big data era.
- To address the limitations of traditional search methods for handling diverse and voluminous scientific literature data.
Main Methods:
- Developed a fast literature retrieval method combining R-tree indexing, optimized for diverse data types and large volumes.
- Utilized an improved k-means clustering algorithm to construct an R-tree clustering model for enhanced retrieval.
- Implemented R-tree indexing to improve the efficiency of searching scientific and technical literature data.
Main Results:
- Experimental results on university science and technology literature datasets demonstrate significant improvements.
- The proposed method shows enhanced efficiency in literature searching.
- The method also achieves improved precision in retrieving relevant scientific literature.
Conclusions:
- The R-tree indexing method offers a viable solution for efficient and precise large-scale scientific literature retrieval.
- This approach effectively addresses the challenges posed by big data in scientific information services.
- The study highlights the potential of advanced indexing techniques for modern knowledge discovery.
More Related Videos
Related Concept Videos
Cluster Sampling Method
12.5K
Appropriate sampling methods ensure that samples are drawn without bias and accurately represent the population. Because measuring the entire population in a study is not practical, researchers use samples to represent the population of interest.
To choose a cluster sample, divide the population into clusters (groups) and then randomly select some of the clusters. All the members from these clusters are in the cluster sample. For example, if you randomly sample four departments from your...
To choose a cluster sample, divide the population into clusters (groups) and then randomly select some of the clusters. All the members from these clusters are in the cluster sample. For example, if you randomly sample four departments from your...
12.5K
Tandem Mass Spectrometry
1.2K
Tandem mass spectrometry is a technique that uses multiple mass analyzers in series to obtain a higher selectivity and signal-to-noise ratio for the analyte. Instruments with multiple analyzers separated by an interaction cell enable secondary fragmentation and selected study of the fragment ions.
Secondary fragmentations occur in the interaction cell and can be induced by various factors. Fragmentation induced by collision with inert gases, such as N2, Ar, He, etc., is called collision-induced...
Secondary fragmentations occur in the interaction cell and can be induced by various factors. Fragmentation induced by collision with inert gases, such as N2, Ar, He, etc., is called collision-induced...
1.2K
Chi-square Analysis
38.7K
The chi-square test is a statistical hypothesis test. It is used to check whether there is a significant difference between an expected value and an observed value. In the context of genetics, it enables us to either accept or reject a hypothesis, based on how much the observed values deviate from the expected values.
The chi-square test was developed by Pearson in 1990.
The first step of performing a Chi-square analysis is to establish a null hypothesis, which assumes that there is no real...
The chi-square test was developed by Pearson in 1990.
The first step of performing a Chi-square analysis is to establish a null hypothesis, which assumes that there is no real...
38.7K
Mass Spectrometry: Overview
5.6K
Mass spectrometry is an analytical technique used to determine the molecular mass and molecular formula of a compound. The basic principle of mass spectrometry is to generate ions from the analyte molecule and measure these ion abundances against their molecular mass. One common type of ionization, known as electrospray ionization or EI, bombards the analyte molecules in the gas phase with high-energy electron beams. The electron beams displace an electron from the molecule and leave...
5.6K
Health Literacy
4.1K
Health literacy is an individual's or a community's capacity to comprehend, receive, read, and use relevant healthcare information and services. The World Health Organization (WHO, 2018) defines health literacy as the cognitive and social skills that determine the ability of individuals to gain access to, understand, and use information in ways that promote and maintain good health. As a result, the WHO helps individuals manage long-term health concerns, participate in preventative...
4.1K
Mass Spectrometers
5.9K
This lesson details the instrumentation of a mass spectrometer—a physical instrument to perform mass spectrometry on analyte molecules and record the characteristic mass spectra. This is achieved via three chief functions:
5.9K

