Mass spectrometry-based proteomics data from thousands of HeLa control samples

Novo Nordisk Foundation
Center for Basic Metabolic Research

Mass spectrometry-based proteomics data from thousands of HeLa control samples

Research output: Contribution to journal › Journal article › Research › peer-review

Documents

Fulltext
Final published version, 2.16 MB, PDF document

Here we provide a curated, large scale, label free mass spectrometry-based proteomics data set derived from HeLa cell lines for general purpose machine learning and analysis. Data access and filtering is a tedious task, which takes up considerable amounts of time for researchers. Therefore we provide machine based metadata for easy selection and overview along the 7,444 raw files and MaxQuant search output. For convenience, we provide three filtered and aggregated development datasets on the protein groups, peptides and precursors level. Next to providing easy to access training data, we provide a SDRF file annotating each raw file with instrument settings allowing automated reprocessing. We encourage others to enlarge this data set by instrument runs of further HeLa samples from different machine types by providing our workflows and analysis scripts.

Original language	English
Article number	112
Journal	Scientific Data
Volume	11
Number of pages	7
ISSN	2052-4463
DOIs	https://doi.org/10.1038/s41597-024-02922-z
Publication status	Published - 2024

Bibliographical note

ID: 381888756

Novo Nordisk Foundation Center for Basic Metabolic Research

Mass spectrometry-based proteomics data from thousands of HeLa control samples

Documents

Bibliographical note