A systematic approach to modeling, capturing, and disseminating proteomics experimental data
- 1 March 2003
- journal article
- research article
- Published by Springer Nature in Nature Biotechnology
- Vol. 21 (3) , 247-254
- https://doi.org/10.1038/nbt0303-247
Abstract
Both the generation and the analysis of proteome data are becoming increasingly widespread, and the field of proteomics is moving incrementally toward high-throughput approaches. Techniques are also increasing in complexity as the relevant technologies evolve. A standard representation of both the methods used and the data generated in proteomics experiments, analogous to that of the MIAME (minimum information about a microarray experiment) guidelines for transcriptomics, and the associated MAGE (microarray gene expression) object model and XML (extensible markup language) implementation, has yet to emerge. This hinders the handling, exchange, and dissemination of proteomics data. Here, we present a UML (unified modeling language) approach to proteomics experimental data, describe XML and SQL (structured query language) implementations of that model, and discuss capture, storage, and dissemination strategies. These make explicit what data might be most usefully captured about proteomics experiments and provide complementary routes toward the implementation of a proteome repository.Keywords
This publication has 14 references indexed in Scilit:
- Empirical Statistical Model To Estimate the Accuracy of Peptide Identifications Made by MS/MS and Database SearchAnalytical Chemistry, 2002
- Minimum information about a microarray experiment (MIAME)—toward standards for microarray dataNature Genetics, 2001
- Bioinformatic assessment of mass spectrometric chemical derivatisation techniques for proteome database searchingProteomics, 2001
- The mouse SWISS-2D PAGE database: a tool for proteomics study of diabetes and obesityProteomics, 2001
- Guilt-by-association goes globalNature, 2000
- The 1999 SWISS-2DPAGE database updateNucleic Acids Research, 2000
- The quest to deduce protein function from sequence: the role of pattern databasesThe International Journal of Biochemistry & Cell Biology, 1999
- Probability-based protein identification by searching sequence databases using mass spectrometry dataElectrophoresis, 1999
- Difference gel electrophoresis. A single gel method for detecting changes in protein extractsElectrophoresis, 1997
- An approach to correlate tandem mass spectral data of peptides with amino acid sequences in a protein databaseJournal of the American Society for Mass Spectrometry, 1994