COG_Functional_Category_Abundances_and_GTDB_Taxonomy
<p><strong>Dataset S1:</strong></p> <p><strong>Individual rows correspond to individual genomes (excluding the top row which are column headers). Columns 1 through 25 correspond to raw abundances for each COG functional category. Column 26 corresponds to the total number of COGs in a genome. Columns 27, 28, 29, 30, 31, 32, and 33, correspond to the GTDB domain, phylum, class, order, family, genus, and species classification, respectively. Column 34 corresponds to the culture-status. Column 35 is the genomes size in base pairs. Column 36 corresponds to the accession number for each genome. Accessions starting with GCF and GCA are from Refseq and Genbank, respectively. Accessions that are numbers only correspond to IMG/G. Column 37 corresponds to the total number of open reading frames in the genome.</strong></p>
ShareScore
40/100
Overall dataset sharing score
Score breakdown
These five areas show where the dataset supports — or may limit — practical reuse.
- Stewardship
- 4
- Harmonization
- 4
- Access
- 20
- Reuse readiness
- 8
- Engagement
- 4