Reflections on Provenance Ontology Encodings

As more data (especially scientific data) is digitized and put on the Web, the importance of tracking and sharing its provenance metadata grows. Besides capturing the annotation properties of data, provenance research also emphasizes interlinking relevant data. Therefore, it is desirable to make provenance metadata easy to access, share, reuse, integrate and reason with. To address these requirements, ontologies can be of use to encode expectations and agreements concerning provenance metadata reuse and integration. The Web is of use to support access and sharing. The Semantic Web, with its languages for representing terms and their descriptions, such as RDFS and OWL, is of use for capturing expectations, agreements, and meaning. We are investigating best practices for providing Semantic Web encodings for provenance ontologies by analyzing a selection of popular Semantic Web provenance ontologies such as Open Provenance Model (OPM), Dublin Core (DC) Terms, and the Proof Markup Language (PML). In this paper, we will highlight a few findings which include: (i) similarities and differences among existing provenance ontologies; (ii) popular approaches used to model provenance concepts and lessons learned from the usage of Semantic Web language features in representing provenance concepts; (iii) expressivity and tractability of representative provenance ontologies. The outcome of our study provides not only guidance to provenance ontology users but also insights to promote better collaborative provenance ontology development and scalable processing of provenance ontologies.

View Publication

Associated Projects

The Inference Web is a Semantic Web based knowledge provenance infrastructure that supports interoperable explanations of sources, assumptions, learned information, and answers as an enabler for trust.

Citation