Skip to main content
zenodoopen

Raw and Intermediate data associated with "Elucidating Human Milk Oligosaccharide biosynthetic genes through network-based multiomics integration"

<p>Data and intermediate structures for&nbsp;&quot;Elucidating Human Milk Oligosaccharide biosynthetic genes through network-based multiomics integration&quot;</p> <p>The project involves the comparison of gene expression data and estimated flux across over 44 million models of HMO biosynthesis. Here we provide intermediate objects to simplify the task of reproducing our calculations.</p> <ul> <li>Raw/Input data (raw.zip); data_HMO.xlsx is the central dataset including HMO concentrations and gene expression data.&nbsp; Other files are .csv exports from data_HMO.xlsx to circumvent a temporary bug in the xlsx reading package. They should not differ from the original file and are provided&nbsp;to facilitate code readability.</li> <li>(models_split.zip) HMO biosynthesis models generated through network generation ( /code/1.flux/A) and enumeration (/code/1.flux/B)</li> <li>(scores_spllit.zip) Correlation (R) and correlation p-values (P) between estimated flux and observed gene expression for corresponding genes (/code/1.flux/C)</li> <li>(big_mat.&lt;bin/desc&gt;.zip) Intermediate matrixes containing gene-linkage scores, models score, model classifications and other information generated/used&nbsp;by the Flux-Expression Comparision code (/code/3.flux_expression)</li> </ul>

ShareScore

28/100

Overall dataset sharing score

Score breakdown

These five areas show where the dataset supports — or may limit — practical reuse.

Stewardship
8
Harmonization
4
Access
12
Reuse readiness
0
Engagement
4

Topics