Demo datasets for the protocol to identify shared transcriptional risks between diseases and compounds predicted to result in mutual benefit
<p>We present a computational protocol (https://github.com/ghbore/protocol-cancer-cvd-similarity), implemented as a Snakemake workflow, that was used in previous works (Gao et al., 2022; Baylis et al., 2023). This protocol allows researchers to identify shared transcriptional processes that drive disease and to screen existing compounds for mutual benefit. The protocol also includes a description of the pharmacovigilance study design used to validate the effect of novel compounds using electronic health records, where applicable. This repository bundles the datasets used in previous works as an example to run through the Snakemake workflow. These datasets include the TCGA cancer dataset, the STARNET and BiKE CVD datasets, and other dependent resources.</p>
ShareScore
32/100
Overall dataset sharing score
Score breakdown
These five areas show where the dataset supports — or may limit — practical reuse.
- Stewardship
- 4
- Harmonization
- 4
- Access
- 16
- Reuse readiness
- 8
- Engagement
- 0