Skip to main content
zenodoopen

COG_Functional_Category_Abundances_and_GTDB_Taxonomy

<p><strong>Dataset S1:</strong></p> <p><strong>Individual rows correspond to individual genomes (excluding the top row which are column headers). Columns 1 through 25 correspond to raw abundances for each COG functional category. Column 26 corresponds to the total number of COGs in a genome. Columns 27, 28, 29, 30, 31, 32, and 33, correspond to the GTDB domain, phylum, class, order, family, genus, and species classification, respectively. Column 34 corresponds to the culture-status. Column 35 is the genomes size in base pairs. Column 36 corresponds to the accession number for each genome. Accessions starting with GCF and GCA are from Refseq and Genbank, respectively. Accessions that are numbers only correspond to IMG/G. Column 37 corresponds to the total number of open reading frames in the genome.</strong></p>

ShareScore

40/100

Overall dataset sharing score

Score breakdown

These five areas show where the dataset supports — or may limit — practical reuse.

Stewardship
4
Harmonization
4
Access
20
Reuse readiness
8
Engagement
4