Lite Kraken/Bracken databases built using UHGG genomes
<p>Kraken/Braken databases for UHGG genomes.</p> <p>HUMAN_3006.tar.gz: Kraken/Bracken database for 3006 high quality species clusters of the UHGG (Beresford-Jones et al., 2022). Database was built from the single highest quality genome for each species cluster (n=3006). Uses the original GTDB v1.3 taxonomy.</p> <p>UHGG_5987_KRAKEN.tar.gz: Kraken/Bracken database for 3006 high quality species clusters of the UHGG (Beresford-Jones et al., 2022). Species clusters are represented by a variable number of high quality genomes (n=5987 in total), selected to maximise represented taxonomic diversity. Uses a custom taxonomy modified from GTDB v2.1 with each species cluster being represented by a species level taxonomic annotation. </p> <p> </p> <p>Methods:</p> <p>Databases built using Kraken v2.1.2 and Bracken v2.6.2. Commands used to build the databases are included below.</p> <p>kraken2-build --build --db Kraken --threads 12</p> <p>bracken-build -d Kraken -k 35 -l 150 -t 12</p>
ShareScore
36/100
Overall dataset sharing score
Score breakdown
These five areas show where the dataset supports — or may limit — practical reuse.
- Stewardship
- 8
- Harmonization
- 4
- Access
- 16
- Reuse readiness
- 8
- Engagement
- 0