Navigation auf zora.uzh.ch

Search ZORA

ZORA (Zurich Open Repository and Archive)

Data provenance: A Cctegorization of existing approaches

Glavic, B; Dittrich, K R (2007). Data provenance: A Cctegorization of existing approaches. In: 12. Fachtagung des GI-Fachbereichs "Datenbanken und Informationssysteme", Aachen, Germany, 7 March 2007 - 9 March 2007. Gesellschaft für Informatik (GI), 227-241.

Abstract

In many application areas like e-science and data-warehousing detailed
information about the origin of data is required. This kind of information is often referred
to as data provenance or data lineage. The provenance of a data item includes
information about the processes and source data items that lead to its creation and
current representation. The diversity of data representation models and application
domains has lead to a number of more or less formal definitions of provenance. Most
of them are limited to a special application domain, data representation model or data
processing facility. Not surprisingly, the associated implementations are also restricted
to some application domain and depend on a special data model. In this paper we give
a survey of data provenance models and prototypes, present a general categorization
scheme for provenance models and use this categorization scheme to study the properties
of the existing approaches. This categorization enables us to distinguish between
different kinds of provenance information and could lead to a better understanding of
provenance in general. Besides the categorization of provenance types, it is important
to include the storage, transformation and query requirements for the different kinds of
provenance information and application domains in our considerations. The analysis
of existing approaches will assist us in revealing open research problems in the area of
data provenance.

Additional indexing

Item Type:Conference or Workshop Item (Paper), refereed, further contribution
Communities & Collections:03 Faculty of Economics > Department of Informatics
Dewey Decimal Classification:000 Computer science, knowledge & systems
Scopus Subject Areas:Physical Sciences > Computer Networks and Communications
Physical Sciences > Information Systems
Uncontrolled Keywords:provenance, survey
Scope:Discipline-based scholarship (basic research)
Language:English
Event End Date:9 March 2007
Deposited On:16 Dec 2009 08:59
Last Modified:06 Mar 2024 13:58
Publisher:Gesellschaft für Informatik (GI)
Series Name:GI-Edition - Lecture notes in informatics (LNI). Proceedings
Number:103
ISBN:978-3-88579-197-3
Additional Information:12. GI-Fachtagung für Datenbanksysteme in Business, Technologie und Web 5. bis 9. März 2007 – Aachen
OA Status:Green
Free access at:Official URL. An embargo period may apply.
Official URL:http://www.btw2007.de/paper/p227.pdf
Related URLs:http://www.gi.de/service/publikationen/lni/gi-edition-lecture-notes-in-informatics-lni-p-103.html
Other Identification Number:merlin-id:384
Download PDF  'Data provenance: A Cctegorization of existing approaches'.
Preview
  • Language: English

Metadata Export

Statistics

Citations

Altmetrics

Downloads

535 downloads since deposited on 16 Dec 2009
24 downloads since 12 months
Detailed statistics

Authors, Affiliations, Collaborations

Similar Publications