Skip to main content
Powered by ShareScore

Find research datasets worth reusing

Search datasets from major research repositories and use ShareScore to quickly assess how well each record supports discovery, access, and reuse.

1

datasets available to search

ShareScore release 0.9.0

Reset

Dataset results

1 result for “Cladocopium”

Learn how ShareScore rates datasets ↗
zenodo28/100

Pocillopora and Cladocopium gene expression levels and Cladocopium SNPs

<p><strong>Pocillopora holobiont gene expression levels</strong></p> <p>The following files contain the genes expression levels of <em>Pocillopora </em>and its<em> Cladocopium </em>photosymbiont for102 <em>Pocillopora </em>coral colonies collected in the framework of Tara Pacific expedition.</p> <p>Pocillopora_MetaT_ReadCount.tab : <em>Pocillopora </em>raw read counts</p> <p>Pocillopora_MetaT_TPM.tab : <em>Pocillopora </em>normalized read counts</p> <p>CladocopiumC1_MetaT_ReadCount.tab : <em>Cladocopium </em>raw read counts</p> <p>CladocopiumC1_MetaT_TPM.tab : <em>Cladocopium </em>normalized read counts</p> <p>Methods: Pocillopora fragments from 102 colonies were processed to extract then sequence RNA. Metatranscriptomic reads (Illumina-generated 150-bp, paired-end) were separately aligned to predicted coding sequences (CDS) of the Pocillopora meandrina coral host reference genome, the CDS of the Cladocopium goreaui genome, and a Durusdinium transcriptome using Burrows&ndash;Wheeler Transform Aligner (BWA-mem, v0.7.15) with the default settings. Host- and symbiont-mapped reads were then sorted and processed using SAMtools v1.10.282 to generate respective bam files. A read was considered a host contig if its sequence aligned to the P. meandrina predicted coding sequence with &ge; 95% of sequence identity and with &ge; 50% of the sequence aligned. Reads aligned to Cladocopium goreaui coding sequences with a cutoff of &ge; 98% of sequence identity over &ge; 80% of the read length were retained as symbiont reads. Reads were further filtered to remove those in which more than 75% of the read length was low complexity or less than 30% was high complexity. Read counts were normalized as transcript per million (TPM).</p> <p><strong>Cladocopium Variants :</strong></p> <p>The following file contains <em>Cladocopium</em> variants<em> </em>called from metatranscriptomic reads of 82 samples aligned on <em>Cladocopium goreaui </em>coding sequences.</p> <p>Cladocopium_FilteredSNPs_4x.vcf.gz : Filtered variants in each sample in vcf format</p> <p>Cladocopium_FilteredSNPs_4x.freq.tab : Alternative allele frequencies of filtered variants in each sample</p> <p>Method: For Pocillopora colonies hosting Cladocopium (84 colonies) we further investigated their population structure using single nucleotide polymorphism (SNP) distributions across their coding sequences. Briefly, we identified a set of transcriptome-wide single nucleotide polymorphisms (SNPs) from metatranscriptomic reads mapped to the Cladocopium goreaui CDS using the Genome Analysis Toolkit tool (GATK, v3.7.0). We followed a modified version of the best practices guide for variant discovery with GATK which included indexing of the genomic reference (picardtools v2.6.0, CreateSequenceDictionary), followed by identification of realignment targets (GATK RealignerTargetCreator) and realignment around detected indels (GATK, IndelRealigner). Variants were called for each colony individually (GATK, HaplotypeCaller) and resulting variant call files (VCFs) were merged into island-specific, multi-sample, cohort files (GATK, CombineGVCFs) before performing joint genotyping across all 11 islands (GATK, GenotypeGVCFs) with polyploidy defined at 1. Joint analysis of multiple samples (i.e., joint genotyping) is recommended for discovery of germline SNPs and indels as it provides information regarding population-wide variance across a cohort of multiple samples. We excluded from this analysis colonies containing a large proportion (&gt;25%) of a second ITS2 profile and potentially affecting the SNP calling. SNPs were filtered using VCFtools (v0.1.12) to include only biallelic SNPs with a quality score &ge; 30 and a coverage &ge; 4</p>

opencc-by-4.0Mar 2022View details →

ScienceDex guides

Understand access before you commit

These curated guides explain access requirements, typical timelines, costs, and reuse considerations for widely used research datasets.

Compare curated datasets

Allen Brain Atlas

Allen Brain Atlas is an Allen Institute collection of brain map atlases, datasets, APIs, and analysis tools covering mouse, human, and non-human primate brain resources.

allen-brain-atlas
neuroscienceopenDocumentation, web resources, and API references are available online.
Last verified 2026-04-30Open record

Annotated Behaviour and Observability Dataset (ABODe)

ABODe is a University of Edinburgh DataShare dataset for behavior classification in group-housed mice using home-cage video, identities, bounding boxes, ground-plate positions, and annotator labels.

abode-home-cage
behavioral-neuroscienceopenThe DataShare record exposes download links for annotations, documentation, license text, and the zipped per-snippet data directory.
Last verified 2026-04-30Open record

DANDI Archive for NWB datasets

DANDI is a BRAIN Initiative archive for publishing and sharing neurophysiology data, including electrophysiology, optophysiology, and behavioral data packaged as NWB and related standards.

dandi-nwb
electrophysiologyopenPublished Dandiset metadata and archive endpoints are available through the production DANDI API.
Last verified 2026-04-30Open record

International Brain Laboratory public data

The International Brain Laboratory public data releases expose standardized mouse decision-making experiments, including Neuropixels recordings, widefield calcium imaging, behavior, and session metadata accessed through the ONE API.

ibl
behavioral-neuroscienceopenPublic sessions can be searched and loaded from the IBL public data server through ONE.
Last verified 2026-04-29Open record

OpenNeuro

OpenNeuro is a free, open platform for sharing neuroimaging datasets, with public search, dataset pages, and download paths for web, S3, DataLad, and the OpenNeuro CLI.

openneuro
neuroscienceopenPublished datasets are available on demand over the internet.
Last verified 2026-04-29Open record