Phables v 0.2.0 benchmarking data and results
<p>This record contains all the benchmarking datasets and results for the Phables manuscript. The datasets used are,</p> <ol> <li>Water samples from Nansi Lake and Dongping Lake in n Shandong Province, China (NCBI BioProject number PRJNA756429), referred to as <strong>Lake Water</strong></li> <li>Soil samples from flooded paddy fields from Hunan Province, China (NCBI BioProject number PRJNA866269), referred to as <strong>Paddy Soil</strong></li> <li>Wastewater virome (NCBI BioProject number PRJNA434744), referred to as <strong>Wastewater</strong></li> <li>Stool samples from patients with IBD and their healthy household controls (NCBI BioProject number PRJEB7772), referred to as <strong>IBD</strong></li> </ol> <p>Each dataset was preprocessed using Hecatomb. The Lake Water dataset was also assembled using MEGAHIT (referred to as <strong>Lake Water - MEGAHIT</strong>) and metaSPAdes (referred to as <strong>Lake Water - metaSPAdes</strong>). All the datasets were run using PHAMB and Phables and the genomes were evaluated using CheckV.</p>
ShareScore
40/100
Overall dataset sharing score
Score breakdown
These five areas show where the dataset supports — or may limit — practical reuse.
- Stewardship
- 8
- Harmonization
- 4
- Access
- 16
- Reuse readiness
- 8
- Engagement
- 4