Data management pipeline for plant phenotyping in a multisite project

20Citations
Citations of this article
65Readers
Mendeley users who have this article in their library.

Abstract

In plant breeding, plants have to be characterised precisely, consistently and rapidly by different people at several field sites within defined time spans. For a meaningful data evaluation and statistical analysis, standardised data storage is required. Data access must be provided on a long-term basis and be independent of organisational barriers without endangering data integrity or intellectual property rights. We discuss the associated technical challenges and demonstrate adequate solutions exemplified in a data management pipeline for a project to identify markers for drought tolerance in potato. This project involves 11 groups from academia and breeding companies, 11 sites and four analytical platforms. Our data warehouse concept combines central data storage in databases and a file server and integrates existing and specialised database solutions for particular data types with new, project-specific databases. The strict use of controlled vocabularies and the application of web-access technologies proved vital to the successful data exchange between diverse institutes and data management concepts and infrastructures. By presenting our data management system and making the software available, we aim to support related phenotyping projects. © 2012 CSIRO.

Cite

CITATION STYLE

APA

Billiau, K., Sprenger, H., Schudoma, C., Walther, D., & Köhl, K. I. (2012). Data management pipeline for plant phenotyping in a multisite project. Functional Plant Biology, 39(11), 948–957. https://doi.org/10.1071/FP12009

Register to see more suggestions

Mendeley helps you to discover research relevant for your work.

Already have an account?

Save time finding and organizing research with Mendeley

Sign up for free