Skip to main content
zenodoopen

Human Myeloma Cell Line (HMCL) NCATS MIPE 4.0 Drug Screen Dataset

<h3>Abstract</h3> <p>Multiple myeloma, a hematopoietic malignancy of terminally differentiated B cells, is the second most common hematological malignancy after leukemia. While patients have benefited from numerous advances in treatment in recent years resulting in significant increases to average survival time following diagnosis, myeloma remains incurable and relapse is common. To help identify novel therapeutic agents with efficacy against the disease and to search for biomarkers associated with differential response to treatment, a large-scale pharmacological screen was performed with 1,912 small molecule compounds tested at 11 doses for 47 human myeloma cell lines (HMCL). Raw and processed versions of the drug screen dataset are provided, as well as supportive information including drug and cell line metadata and high-level characterization of the most salient features of each. The dataset is publicly available at Zenodo and the workflow code used for data processing and generation of supporting figures and tables can be found at https://github.com/khughitt/hmcl-drug-screen-pipeline.</p> <h3>Overview</h3> <p>The dataset shared here contains the raw and processed versions of a high throughput small molecule&nbsp;screening dataset, as as supporting tables and figures.</p> <p>The computational pipeline used to generate the dataset is available at:&nbsp;<a href="https://github.com/khughitt/hmcl-drug-screen-pipeline">https://github.com/khughitt/hmcl-drug-screen-pipeline</a></p> <h3>Data organization</h3> <p>The .zip file available on Zenodo contains the complete output from the computational pipeline&nbsp;linked to above.</p> <p>Data and figures are organized into subfolders relating to the various pipeline steps.</p> <p>Files &amp; folders:</p> <table> <tbody> <tr> <td><strong>drugs/</strong></td> <td>Drug curves, cell line x drug AC-50 matrix, and drug metadata tables.</td> </tr> <tr> <td><strong>manuscript/</strong></td> <td>Figures and tables generated by the pipeline for the manuscript, including the harmonized cell line metadata table ("table1.tsv")</td> </tr> <tr> <td><strong>mutations/</strong></td> <td> <p>Cell line mutation predictions with identifiers harmonized to match those used in the drug screen (source: <a href="https://www.keatslab.org/data-repository">Keats Lab</a>)</p> </td> </tr> <tr> <td><strong>plates/</strong></td> <td> <p>Drug screen viability measurements at different levels of processing, arranged with one column per plate ("raw.tsv", "raw_filtered.tsv", "normed.tsv", "background_adjusted.tsv"), plate-level metadata ("metadata.tsv"), the compued background plate ("background.tsv"), and plate images used for QA.</p> </td> </tr> <tr> <td><strong>raw/</strong></td> <td>Raw version of the drug screen dataset, including the data for cell lines which were filtered out in the pipeline due to quality purposes</td> </tr> <tr> <td><strong>datapackage.yml</strong></td> <td> <p>Metadata for the provided data resources including sha256 checksums, human-readible descriptions, column types, etc. (see: <a href="datapackage.org">https://datapackage.org/</a>)</p> </td> </tr> </tbody> </table> <p>&nbsp;</p>

ShareScore

28/100

Overall dataset sharing score

Score breakdown

These five areas show where the dataset supports — or may limit — practical reuse.

Stewardship
4
Harmonization
4
Access
16
Reuse readiness
0
Engagement
4