Storage

Efficient Querying of Distributed Provenance Stores

Free registration required

Executive Summary

Current projects that automate the collection of provenance information use a centralized architecture for managing the resulting metadata - that is, provenance is gathered at remote hosts and submitted to a central provenance management service. In contrast, the authors are developing a completely decentralized system with each computer maintaining the authoritative repository of the provenance gathered on it. Their model has several advantages, such as scaling to large amounts of metadata generation, providing low-latency access to provenance metadata about local data, avoiding the need for synchronization with a central service after operating while disconnected from the network, and letting users retain control over their data provenance records.

  • Format: PDF
  • Size: 193.6 KB