Skip to main content
zenodoopen

The PODER cross-ancestry long-read RNA-seq dataset

<p>Table legends:</p> <ul> <li>column_descriptions.docx: Extended description of columns in selected tables.</li> <li>00_sample_metadata: Sample and experimental metadata pertaining to each LR-RNA-seq sample.</li> <li>01_uma_gtf: Unfiltered merged annotation (UMA) GTF for all spliced intron chains discovered by each of the tools.</li> <li>02_uma_mt: Metadata table for SQANTI QC, Recount3 support, protein prediction, and other relevant characteristics for each transcript in the UMA annotation used to filter transcripts.</li> <li>03_poder_gtf: Final PODER GTF, including predicted CDSs for novel transcripts and annotated CDSs for known transcripts.</li> <li>04_poder_mt: Metadata table for SQANTI QC, Recount3 support, protein prediction, and other relevant characteristics for each transcript in PODER.</li> <li>06_poder_t_counts: PODER transcript counts in each sample as quantified by lr-kallisto.</li> <li>07_poder_g_counts: PODER gene counts in each sample as quantified by lr-kallisto.</li> <li>08_mage_t_counts_poder: Transcript counts for the MAGE RNA-seq dataset computed using the PODER annotation with kallisto.</li> <li>09_mage_t_counts_gencode: Transcript counts for the MAGE RNA-seq dataset computed using the GENCODE annotation with kallisto.</li> <li>10_mage_t_counts_enh: Transcript counts for the MAGE RNA-seq dataset computed using the Enhanced GENCODE annotation with kallisto.</li> <li>11_astu: Allele-specific transcript usage results.</li> <li>12_ase: Allele-specific expression results.</li> <li>13_mage_tau: Tau values for Enhanced GENCODE transcripts computed on the MAGE RNA-seq dataset.</li> <li>14_gwas_enrichments: GWAS enrichment results for ASTU genes overlapping GWAS genes.</li> <li>16_inter_catalog_overlap: Boolean detection of transcripts across PODER, annotations (GENCODE, RefSeq) and other RNA-seq / LR-RNA-seq-derived transcript catalogs (CHESS, GTEx, ENCODE4).</li> <li>17_personalized_hg38: Isoforms and their intron chains detected using personalized-GRCh38s.</li> <li>18_enh_gencode: Enhanced GENCODE GTF with novel transcripts from PODER added to all annotated transcripts from GENCODE v47</li> </ul>

ShareScore

40/100

Overall dataset sharing score

Score breakdown

These five areas show where the dataset supports — or may limit — practical reuse.

Stewardship
8
Harmonization
4
Access
20
Reuse readiness
8
Engagement
0