Data curation + process curation = data integration + science

Carole Goble, Robert Stevens, Duncan Hull, Katy Wolstencroft, Rodrigo Lopez

    Research output: Contribution to journalArticlepeer-review

    Abstract

    In bioinformatics, we are familiar with the idea of curated data as a prerequisite for data integration. We neglect, often to our cost, the curation and cataloguing of the processes that we use to integrate and analyse our data. Programmatic access to services, for data and processes, means that compositions of services can be made that represent the in silico experiments or processes that bioinformaticians perform. Data integration through workflows depends on being able to know what services exist and where to find those services. The large number of services and the operations they perform, their arbitrary naming and lack of documentation, however, mean that they can be difficult to use. The workflows themselves are composite processes that could be pooled and reused but only if they too can be found and understood. Thus appropriate curation, including semantic mark-up, would enable processes to be found, maintained and consequently used more easily. This broader view on semantic annotation is vital for full data integration that is necessary for the modern scientific analyses in biology. This article will brief the community on the current state of the art and the current challenges for process curation, both within and without the Life Sciences. © The Author 2008. Published by Oxford University Press.
    Original languageEnglish
    Pages (from-to)506-517
    Number of pages11
    JournalBriefings in Bioinformatics
    Volume9
    Issue number6
    DOIs
    Publication statusPublished - 2008

    Keywords

    • Curation
    • Metadata
    • Ontology
    • Processes
    • Semantic annotation
    • Services
    • Workflow

    Fingerprint

    Dive into the research topics of 'Data curation + process curation = data integration + science'. Together they form a unique fingerprint.

    Cite this