An identity crisis in the life sciences

Zhao, Jun and Goble, Carole and Stevens, Robert (2006) An identity crisis in the life sciences. In: Provenance and Annotation of Data. Lecture Notes in Computer Science . Springer, Berlin, pp. 254-269. ISBN 9783540463023

Full text not available from this repository.


myGrid is an e-Science project assisting life scientists to build workflows that gather data from distributed, autonomous, replicated and heterogeneous resources. The provenance logs of workflow executions are recorded as RDF graphs. The log of one workflow run is used to trace the history of its execution process. However, by aggregating provenance logs of many workflow runs, one may gather the provenance of a common data product shared in multiple derivation paths. A successful aggregation relies on accurate and universal identification of each data product. The nature of bioinformatics data and services, however, makes this difficult. We describe the identity problem in bioinformatics data, and present a protocol for managing identity co-references and allocating identity to gathered and computed data products. The ability to overcome this problem means that the provenance of workflows in bioinformatics and other domains can be exploited to enhance the practice of e-Science.

Item Type:
Contribution in Book/Report/Proceedings
ID Code:
Deposited By:
Deposited On:
21 Aug 2014 08:17
Last Modified:
18 Sep 2023 02:33