Find research datasets worth reusing
Search datasets from major research repositories and use ShareScore to quickly assess how well each record supports discovery, access, and reuse.
10
datasets available to search
ShareScore release 0.9.0
Dataset results
10 results for “concordance analysis”
Meta-analysis of diurnal transcriptomics reveals strong patterns of concordance and discordance in mouse liver: processed data
<p>The accumulation of public transcriptomic timeseries data enables robust meta-analyses that were not possible until recently. To assess the consistency of biological rhythms across studies, 43 public mouse liver tissue timeseries totaling 805 RNA-seq samples were obtained and analyzed. Only the control groups of each study were included, to create comparable data. Technical factors in RNA-seq library preparation were the largest contributors to transcriptome-level differences, beyond biological or experiment-specific factors such as lighting conditions. Core clock genes were remarkably consistent in phase across all studies, while phase distributions of other periodic genes were generally less consistent. Overlap of genes identified as rhythmic across studies was generally low, with around 50% between some of the highest sample count studies. Distributions of phases of significant genes were remarkably inconsistent across studies, but genes consistently identified as rhythmic clustered near ZT0 and ZT12 in acrophase. Data was integrated across studies in a JIVE analysis, which showed that the top two components of joint within-study variation are determined by time of day. A shape-invariant model with random effects was fit to the genes to identify the underlying shape of the rhythms, consistent across all studies. This revealed the extent of asymmetric and multimodal genes.<br> <br> This supplemental file provides preprocessed RNA-seq quantifications of all reviewed datasets, as well as results of multiple analyses.</p>
Data from: Concordance analysis in mitogenomic phylogenetics
Open the record for dataset details and reuse information.
Patterns of age-related water diffusion changes in human brain by concordance and discordance analysis.
Open the record for dataset details and reuse information.
Proteogenomic analysis of psoriasis reveals discordant and concordant changes in mRNA and protein abundance
GEO Series GSE67785. Homo sapiens. 28 samples. Type: Expression profiling by high throughput sequencing.
Simultaneous profiling of sexually transmitted bacterial pathogens, microbiome, and concordant host response in cervical samples using whole transcriptome sequencing analysis
GEO Series GSE120192. Homo sapiens. 20 samples. Type: Expression profiling by high throughput sequencing.
Epigenome-wide analysis of DNA methylation in lung tissue shows concordance with blood studies and identifies tobacco smoke-inducible enhancers
GEO Series GSE94986. Homo sapiens. 8 samples. Type: Genome binding/occupancy profiling by high throughput sequencing; Methylation profiling by high throughput sequencing.
Epigenome-wide analysis of DNA methylation in lung tissue shows concordance with blood studies and identifies tobacco smoke-inducible enhancers
GEO Series GSE69770. Homo sapiens. 11 samples. Type: Expression profiling by high throughput sequencing.
Data from: Analysis of phylogenomic datasets reveals conflict, concordance, and gene duplications with examples from animals and plants
Background: The use of transcriptomic and genomic datasets for phylogenetic reconstruction has become increasingly common as researchers attempt to resolve recalcitrant nodes with increasing amounts of data. The large size and complexity of these datasets introduce significant phylogenetic noise and conflict into subsequent analyses. The sources of conflict may include hybridization, incomplete lineage sorting, or horizontal gene transfer, and may vary across the phylogeny. For phylogenetic analysis, this noise and conflict has been accommodated in one of several ways: by binning gene regions into subsets to isolate consistent phylogenetic signal; by using gene-tree methods for reconstruction, where conflict is presumed to be explained by incomplete lineage sorting (ILS); or through concatenation, where noise is presumed to be the dominant source of conflict. The results provided herein emphasize that analysis of individual homologous gene regions can greatly improve our understanding of the underlying conflict within these datasets. Results: Here we examined two published transcriptomic datasets, the angiosperm group Caryophyllales and the aculeate Hymenoptera, for the presence of conflict, concordance, and gene duplications in individual homologs across the phylogeny. We found significant conflict throughout the phylogeny in both datasets and in particular along the backbone. While some nodes in each phylogeny showed patterns of conflict similar to what might be expected with ILS alone, the backbone nodes also exhibited low levels of phylogenetic signal. In addition, certain nodes, especially in the Caryophyllales, had highly elevated levels of strongly supported conflict that cannot be explained by ILS alone. Conclusion: This study demonstrates that phylogenetic signal is highly variable in phylogenomic data sampled across related species and poses challenges when conducting species tree analyses on large genomic and transcriptomic datasets. Further insight into the conflict and processes underlying these complex datasets is necessary to improve and develop adequate models for sequence analysis and downstream applications. To aid this effort, we developed the open source software phyparts (https://bitbucket.org/blackrim/phyparts), which calculates unique, conflicting, and concordant bipartitions, maps gene duplications, and outputs summary statistics such as internode certainy (ICA) scores and node-specific counts of gene duplications.
Data from: Analysis of phylogenomic datasets reveals conflict, concordance, and gene duplications with examples from animals and plants
Open the record for dataset details and reuse information.
Unveiling A Hidden Burden: Exploring Sarcopenia in Hospitalized Older Patients through Concordance and Cluster Analysis
Open the record for dataset details and reuse information.
ScienceDex guides
Understand access before you commit
These curated guides explain access requirements, typical timelines, costs, and reuse considerations for widely used research datasets.
Allen Brain Atlas
Allen Brain Atlas is an Allen Institute collection of brain map atlases, datasets, APIs, and analysis tools covering mouse, human, and non-human primate brain resources.
Annotated Behaviour and Observability Dataset (ABODe)
ABODe is a University of Edinburgh DataShare dataset for behavior classification in group-housed mice using home-cage video, identities, bounding boxes, ground-plate positions, and annotator labels.
DANDI Archive for NWB datasets
DANDI is a BRAIN Initiative archive for publishing and sharing neurophysiology data, including electrophysiology, optophysiology, and behavioral data packaged as NWB and related standards.
International Brain Laboratory public data
The International Brain Laboratory public data releases expose standardized mouse decision-making experiments, including Neuropixels recordings, widefield calcium imaging, behavior, and session metadata accessed through the ONE API.
OpenNeuro
OpenNeuro is a free, open platform for sharing neuroimaging datasets, with public search, dataset pages, and download paths for web, S3, DataLad, and the OpenNeuro CLI.