Efficient Querying of Distributed Provenance Stores

Current projects that automate the collection of provenance information use a centralized architecture for managing the resulting metadata - that is, provenance is gathered at remote hosts and submitted to a central provenance management service. In contrast, the authors are developing a completely decentralized system with each computer maintaining the authoritative repository of the provenance gathered on it. Their model has several advantages, such as scaling to large amounts of metadata generation, providing low-latency access to provenance metadata about local data, avoiding the need for synchronization with a central service after operating while disconnected from the network, and letting users retain control over their data provenance records.

Provided by: Association for Computing Machinery Topic: Storage Date Added: Jun 2010 Format: PDF

Download Now

Find By Topic