Skip to main content
zenodoopen

Phables v 0.2.0 benchmarking data and results

<p>This record contains all the benchmarking&nbsp;datasets and results for the&nbsp;Phables manuscript. The datasets used are,</p> <ol> <li>Water samples from Nansi Lake and Dongping Lake in n Shandong Province, China (NCBI BioProject number PRJNA756429), referred to as <strong>Lake Water</strong></li> <li>Soil samples from flooded paddy fields from Hunan Province, China (NCBI BioProject number PRJNA866269), referred to as <strong>Paddy Soil</strong></li> <li>Wastewater virome (NCBI BioProject number PRJNA434744), referred to as <strong>Wastewater</strong></li> <li>Stool samples from patients with IBD and their healthy household controls (NCBI BioProject number PRJEB7772), referred to as <strong>IBD</strong></li> </ol> <p>Each dataset was preprocessed using Hecatomb. The Lake Water dataset was also assembled using MEGAHIT (referred to as <strong>Lake Water - MEGAHIT</strong>) and metaSPAdes (referred to as <strong>Lake Water - metaSPAdes</strong>). All the datasets were run using PHAMB and Phables and the genomes were evaluated using CheckV.</p>

ShareScore

40/100

Overall dataset sharing score

Score breakdown

These five areas show where the dataset supports — or may limit — practical reuse.

Stewardship
8
Harmonization
4
Access
16
Reuse readiness
8
Engagement
4

Topics