Related Experiment Video
Updated: May 30, 2025

A Knowledge Graph Approach to Elucidate the Role of Organellar Pathways in Disease via Biomedical Reports
Published on: October 13, 2023
An Automated Approach for Domain-Specific Knowledge Graph Generation─Graph Measures and Characterization.
Connor O'Ryan1, Kevin D Hayes1, Francis G VanGessel2
1Center for Engineering Concepts Development, Department of Mechanical Engineering, University of Maryland, College Park, Maryland 20742, United States.
This study introduces a natural language processing (NLP) method to create knowledge graphs from scientific texts, enabling better information extraction from vast chemical literature and identifying linguistic trends in synthesis.
Area of Science:
- Computational Chemistry
- Natural Language Processing
- Data Science
Background:
- The exponential growth of scientific publications necessitates automated methods for information extraction.
- Existing knowledge graph extraction methods are often domain-specific and limited.
- Bridging the gap between AI advancements and scientific literature analysis is crucial.
Purpose of the Study:
- To develop a novel natural language processing (NLP) approach for extracting knowledge graphs from technical documents.
- To create a semantically structured network (SSN) from synthetic chemistry patents.
- To characterize the resulting knowledge graph for linguistic and trend analysis.
Main Methods:
- Developed a natural language processing (NLP) model for knowledge graph extraction.
- Applied the model to approximately 100,000 full-length synthetic chemistry patents.
- Performed graph characterization using network motif structures, assortativity, and eigenvector centrality.
Main Results:
- Successfully extracted a semantically structured network (SSN) from chemical patents.
- Identified linguistic patterns in chemical reaction discourse, including common solvents and compound naming.
- Observed power-law trends in larger text corpora, indicating scalability.
Conclusions:
- The developed NLP approach provides a robust method for knowledge graph extraction in specialized domains.
- Quantitative characterization of knowledge graphs aids in understanding scientific discourse and validating large datasets.
- This work facilitates deeper insights into chemical synthesis literature and enables cross-domain knowledge graph comparisons.
Related Concept Videos
Model Approaches for Pharmacokinetic Data: Distributed Parameter Models
The distributed parameter models are specifically designed to account for variations and differences in some drug classes. This model is particularly useful for assessing regional concentrations of anticancer or...
Heuristics
People often rely on heuristics when faced with an overload of information, limited time, low importance of the decision, limited information, or when a heuristic readily comes to mind. For...
Multiple Bar Graph
Each bar or column in the multiple bar graph represents a data value. These graphs are used primarily in interrelating two or more sets of data. The categories of different kinds of data are listed along the horizontal or x-axis, whereas...
Rate-Determining Steps
In a multistep reaction mechanism, one of the elementary steps progresses significantly slower than the others. This slowest step is called the rate-limiting step (or rate-determining step). A reaction cannot proceed faster than its slowest step, and hence, the rate-determining step limits the overall reaction rate.
The concept of rate-determining step can be understood from the analogy of a 4-lane freeway with a short-stretch of traffic-bottleneck caused due to...
Mechanistic Models: Compartment Models in Algorithms for Numerical Problem Solving
In individual population analyses, different algorithms are employed, such as Cauchy's method, which uses a...
Analysis Methods of Pharmacokinetic Data: Model and Model-Independent Approaches
The model approach uses mathematical models to describe changes in drug concentration over time. Pharmacokinetic models help characterize drug behavior in patients, predict drug concentration in the body fluids, calculate optimum dosage regimens, and evaluate the risk of toxicity. However, ensuring that the model fits the experimental data accurately...

