Design of a data model for developing laboratory information management and analysis systems for protein production

Proteins. 2005 Feb 1;58(2):278-84. doi: 10.1002/prot.20303.

Abstract

Data management has emerged as one of the central issues in the high-throughput processes of taking a protein target sequence through to a protein sample. To simplify this task, and following extensive consultation with the international structural genomics community, we describe here a model of the data related to protein production. The model is suitable for both large and small facilities for use in tracking samples, experiments, and results through the many procedures involved. The model is described in Unified Modeling Language (UML). In addition, we present relational database schemas derived from the UML. These relational schemas are already in use in a number of data management projects.

Publication types

  • Research Support, Non-U.S. Gov't
  • Research Support, U.S. Gov't, Non-P.H.S.
  • Research Support, U.S. Gov't, P.H.S.

MeSH terms

  • Algorithms
  • Amino Acid Sequence
  • Data Interpretation, Statistical
  • Databases, Protein
  • Genomics / methods*
  • Internet
  • Models, Biological
  • Programming Languages
  • Protein Engineering / methods*
  • Proteins / chemistry*
  • Proteomics / methods*
  • Research
  • Software
  • Software Design
  • Systems Biology
  • Unified Medical Language System

Substances

  • Proteins