Abstract
The amount of public proteomics data is rapidly increasing but there is no standardized format to describe the sample metadata and their relationship with the dataset files in a way that fully supports their understanding or reanalysis. Here we propose to develop the transcriptomics data format MAGE-TAB into a standard representation for proteomics sample metadata. We implement MAGE-TAB-Proteomics in a crowdsourcing project to manually curate over 200 public datasets. We also describe tools and libraries to validate and submit sample metadata-related information to the PRIDE repository. We expect that these developments will improve the reproducibility and facilitate the reanalysis and integration of public proteomics datasets.
| Original language | English |
|---|---|
| Article number | 5854 |
| Journal | Nature Communications |
| Volume | 12 |
| Issue number | 1 |
| DOIs | |
| State | Published - Dec 1 2021 |
Fingerprint
Dive into the research topics of 'A proteomics sample metadata representation for multiomics integration and big data analysis'. Together they form a unique fingerprint.Cite this
- APA
- Author
- BIBTEX
- Harvard
- Standard
- RIS
- Vancouver