Related Experiment Video
Updated: Jul 2, 2026

05:47
Evidence-based Knowledge Synthesis and Hypothesis Validation: Navigating Biomedical Knowledge Bases via Explainable AI and Agentic Systems
Published on: June 13, 2025
Knowledge Engineering for Open Science: Building and Deploying Knowledge Bases for Metadata Standards
Mark A Musen1, Martin J O'Connor1, Josef Hardi1
1Division of Computational Medicine, Stanford University School of Medicine, Stanford, California, USA.
Summary
Scientists need standardized metadata for FAIR data. The Center for Expanded Data Annotation and Retrieval (CEDAR) uses templates to encode metadata standards, improving data discoverability and reusability.
Area of Science:
- Scientific data management
- Open science initiatives
- Metadata standards development
Background:
- Achieving FAIR data principles (Findable, Accessible, Interoperable, Reusable) requires rich, standardized metadata.
- Lack of standardized metadata hinders data understanding and reuse across scientific disciplines.
- Existing metadata standards can be difficult for researchers and curators to implement.
Purpose of the Study:
- To introduce the Center for Expanded Data Annotation and Retrieval (CEDAR) technology for creating and applying metadata standards.
- To demonstrate how CEDAR templates facilitate community-driven metadata standardization.
- To promote the adoption of standardized metadata for enhanced data sharing and open science.
Main Methods:
- Development of CEDAR technology to encode metadata standards as templates.
- Templates enumerate experimental attributes and link them to ontologies or controlled vocabularies.
- Application of CEDAR templates to standardize metadata for scientific consortia and data annotation systems.
Main Results:
- CEDAR templates capture community metadata preferences, enabling standardized data description.
- Templates have been used to standardize metadata for various scientific consortia.
- CEDAR-based systems facilitate metadata acquisition (e.g., via web forms, spreadsheets) and ensure standard adherence.
Conclusions:
- CEDAR provides a mechanism for scientific communities to establish and deploy shared metadata standards.
- The technology encodes community knowledge in a symbolic form for application in intelligent systems.
- CEDAR promotes open science by improving data findability, accessibility, interoperability, and reusability.
Related Concept Videos
Genome Annotation and Assembly
The genome refers to all of the genetic material in an organism. It can range from a few million base pairs in microbial cells to several billion base pairs in many eukaryotic organisms. Genome assembly refers to the process of taking the DNA sequencing data and putting it all back together in a correct order to create a close representation of the original genome. This is followed by the identification of functional elements on the newly assembled genome, a process called genome annotation.
Levels of Use of a GIS
Geographic Information Systems (GIS) operate across three levels of application, each representing an increasing degree of complexity: data management, analysis, and prediction. These levels reflect the expanding functionality and versatility of GIS technology in handling spatial data for diverse purposes.Data ManagementAt its foundational level, GIS serves as a tool for data management, enabling the input, storage, retrieval, and organization of spatial data. This level is often employed in...
Methods of Documentation VII: EMR
Electronic Medical Records (EMRs) primarily center around electronically documenting patients' health information within a single healthcare organization or practice. They contain essential clinical data related to a patient's medical history, diagnoses, medications, treatment plans, lab results, and other pertinent information relevant to the specific encounter or episode of care. EMRs are designed to streamline documentation and workflow processes within individual healthcare settings,...