Data curation + process curation=data integration + science.

Brief Bioinform

School of Computer Science, University of Manchester, Oxford Road, Manchester, M13 9PL, UK.

Published: November 2008

In bioinformatics, we are familiar with the idea of curated data as a prerequisite for data integration. We neglect, often to our cost, the curation and cataloguing of the processes that we use to integrate and analyse our data. Programmatic access to services, for data and processes, means that compositions of services can be made that represent the in silico experiments or processes that bioinformaticians perform. Data integration through workflows depends on being able to know what services exist and where to find those services. The large number of services and the operations they perform, their arbitrary naming and lack of documentation, however, mean that they can be difficult to use. The workflows themselves are composite processes that could be pooled and reused but only if they too can be found and understood. Thus appropriate curation, including semantic mark-up, would enable processes to be found, maintained and consequently used more easily. This broader view on semantic annotation is vital for full data integration that is necessary for the modern scientific analyses in biology. This article will brief the community on the current state of the art and the current challenges for process curation, both within and without the Life Sciences.

Download full-text PDF

Source
http://dx.doi.org/10.1093/bib/bbn034DOI Listing

Publication Analysis

Top Keywords

data integration
12
data
7
processes
5
services
5
data curation
4
curation process
4
process curation=data
4
integration
4
curation=data integration
4
integration science
4

Similar Publications

Want AI Summaries of new PubMed Abstracts delivered to your In-box?

Enter search terms and have AI summaries delivered each week - change queries or unsubscribe any time!