Related Experiment Video
Updated: May 5, 2026

Development of Compendium for Esophageal Squamous Cell Carcinoma
Published on: April 12, 2024
PAV ontology: provenance, authoring and versioning
Paolo Ciccarese1, Stian Soiland-Reyes, Khalid Belhajjame
1Department of Neurology, Massachusetts General Hospital, 55 Fruit Street, Boston, MA 02114, USA. paolo.ciccarese@gmail.com.
The Provenance, Authoring and Versioning ontology (PAV) offers a lightweight solution for tracking digital scientific content provenance. PAV enhances trust by distinguishing agent roles and resource versions, improving upon existing models like W3C PROV-O.
Area of Science:
- Digital scientific content management
- Knowledge representation and reasoning
Background:
- Establishing trust in scientific content relies heavily on provenance tracking.
- Existing general-purpose vocabularies like Dublin Core Terms (DC Terms) and W3C Provenance Ontology (PROV-O) lack specific classes for agent roles (author, contributor, curator) in digital artifact manipulation.
- PROV-O offers a basic methodology for web resource authoring and versioning but requires extensions for detailed role identification.
Purpose of the Study:
- To introduce the lightweight Provenance, Authoring and Versioning ontology (PAV) for capturing essential descriptions of web resource provenance, authoring, and versioning.
- To address the need for specific classes and properties to distinguish between various agent roles and resource transformations in scientific digital content.
- To enhance interoperability by providing mappings to existing ontologies like W3C PROV-O.
Main Methods:
- Designed PAV with a focus on being lightweight and pragmatic, incorporating requirements from real-world projects.
- Included only terms demonstrated to be useful in existing applications.
- Recommended terms from existing ontologies where plausible to maintain compactness and encourage reuse.
Main Results:
- The Provenance, Authoring and Versioning ontology (PAV) has been developed to capture essential provenance, authoring, and versioning information for digital scientific content.
- PAV distinguishes between content contributors, authors, curators, and representation creators, alongside the provenance of originating resources.
- Five projects have adopted PAV, demonstrating its practical application and providing concrete examples of its usage.
- Mappings have been developed to show how PAV extends W3C PROV-O, enhancing interoperability.
Conclusions:
- PAV provides a focused and lightweight solution for tracking the provenance, authoring, and versioning of digital scientific content.
- PAV offers a clear distinction between agent roles and resource provenance, addressing limitations in existing general-purpose ontologies.
- Comparisons with related approaches (PRV, DC Terms, BIBFRAME) highlight PAV's strengths in pragmatic applicability and specific role distinctions.
- SKOS mappings facilitate alignment between PAV and DC Terms, further promoting interoperability.
Related Concept Videos
Methods of Documentation II: POMR
Methods of Documentation I: Source-Oriented Records
In an SOR, each discipline involved in patient care maintains a separate medical record section. This record-keeping method enables easy tracking of patient progress and ensures healthcare staff have access to up-to-date information.
Key Attributes include the following:
Methods of Documentation III: PIE
Pedigree Analysis
Pedigree Analysis
Methods of Documentation V: CBE
In CBE, healthcare professionals establish predefined standards of practice that define what constitutes...

