Skip to main content
Powered by ShareScore

Find research datasets worth reusing

Search datasets from major research repositories and use ShareScore to quickly assess how well each record supports discovery, access, and reuse.

49

datasets available to search

ShareScore release 0.7.1

Reset

Dataset results

49 results for “pseudogenization”

Learn how ShareScore rates datasets ↗
geo24/100

The Pseudogene RPS27AP5 Reveals Novel Ubiquitin and Ribosomal Protein Variants Involved in Specialised Ribosomal Functions.

GEO Series GSE254888. Homo sapiens. 6 samples. Type: Expression profiling by high throughput sequencing.

openGEO-OpenSep 2025View details →
geo24/100

Positive natural selection of N6-methyladenosine on the RNAs of processed pseudogenes

GEO Series GSE172219. Homo sapiens. 18 samples. Type: Expression profiling by high throughput sequencing.

openGEO-OpenMay 2021View details →
geo24/100

Pseudogenes limit the identification of novel common transcripts generated by their parent genes

GEO Series GSE215459. Homo sapiens. 2 samples. Type: Expression profiling by high throughput sequencing.

openGEO-OpenOct 2022View details →
geo24/100

Exhaustive profiling in Arabidopsis reveals abundant polysome-associated 24-nt small RNAs including hitherto undescribed AGO5-associated pseudogene-derived siRNAs (psiRNAs) [sRNA]

GEO Series GSE99827. Arabidopsis thaliana. 6 samples. Type: Non-coding RNA profiling by high throughput sequencing.

openGEO-OpenDec 2017View details →
geo24/100

Systematic functional interrogation of human pseudogenes

GEO Series GSE155510. Homo sapiens. 12 samples. Type: Other.

openGEO-OpenAug 2021View details →
geo24/100

Pseudogene INTS6P1 regulates its cognate gene INTS6 through competitive binding of miR-17-5p in hepatocellular carcinoma

GEO Series GSE64633. Homo sapiens. 12 samples. Type: Expression profiling by array; Non-coding RNA profiling by array.

openGEO-OpenJan 2015View details →
geo24/100

Long-read cDNA sequencing identifies functional pseudogenes in the human transcriptome

GEO Series GSE160383. Homo sapiens. 19 samples. Type: Expression profiling by high throughput sequencing.

openGEO-OpenApr 2021View details →
geo24/100

Knockdown of pseudogene lncRNA UBE2CP3 in gastric cancer cells

GEO Series GSE163813. Homo sapiens. 9 samples. Type: Expression profiling by high throughput sequencing.

openGEO-OpenMay 2021View details →
geo24/100

A FTH1 gene:pseudogene:miRNA network regulates tumorigenesis in prostate cancer

GEO Series GSE101837. Homo sapiens. 24 samples. Type: Expression profiling by array.

openGEO-OpenJul 2017View details →
geo24/100

A pseudogene lncRNA at the interface of inflammation and anti-inflammatory therapeutics

GEO Series GSE47494. Mus musculus. 6 samples. Type: Expression profiling by high throughput sequencing; Genome binding/occupancy profiling by high throughput sequencing.

openGEO-OpenAug 2013View details →
dryad24/100

Data from: Annotation of pseudogenic gene segments by massively parallel sequencing of rearranged lymphocyte receptor loci

Background: The adaptive immune system generates a remarkable range of antigen-specific T-cell receptors (TCRs), allowing the recognition of a diverse set of antigens. Most of this diversity is encoded in the complementarity determining region 3 (CDR3) of the β chain of the αβ TCR, which is generated by somatic recombination of noncontiguous variable (V), diversity (D), and joining (J) gene segments. Deletion and non-templated insertion of nucleotides at the D-J and V-DJ junctions further increases diversity. Many of these gene segments are annotated as non-functional owing to defects in their primary sequence, the absence of motifs necessary for rearrangement, or chromosomal locations outside the TCR locus. Methods: We sought to utilize a novel method, based on high-throughput sequencing of rearranged TCR genes in a large cohort of individuals, to evaluate the use of functional and non-functional alleles. We amplified and sequenced genomic DNA from the peripheral blood of 587 healthy volunteers using a multiplexed polymerase chain reaction assay that targets the variable region of the rearranged TCRβ locus, and we determined the presence and the proportion of productive rearrangements for each TCRβ V gene segment in each individual. We then used this information to annotate the functional status of TCRβ V gene segments in this cohort. Results: For most TCRβ V gene segments, our method agrees with previously reported functional annotations. However, we identified novel non-functional alleles for several gene segments, some of which were used exclusively in our cohort to the detriment of reported functional alleles. We also saw that some gene segments reported to have both functional and non-functional alleles consistently behaved in our cohort as either functional or non-functional, suggesting that some reported alleles were not present in the population studied. Conclusions: In this proof-of-principle study, we used high-throughput sequencing of the TCRβ locus of a large cohort of healthy volunteers to evaluate the use of functional and non-functional alleles of individual TCRβ V gene segments. With some modifications, our method has the potential to be extended to gene segments in the α, γ, and δ TCR loci, as well as the genes encoding for B-cell receptor chains.

opencc-zeroDec 2014View details →
ClinicalTrials.gov24/100

Role of Pseudogene in Incontinentia Pigmenti, and Its Potential Treatment

ClinicalTrials.gov study NCT00976586. IPD Sharing: Not stated. Countries: 1. Publications: 0.

restrictedIPD-UNDECIDEDFeb 2026View details →
dryad24/100

Data from: Annotation of pseudogenic gene segments by massively parallel sequencing of rearranged lymphocyte receptor loci

Open the record for dataset details and reuse information.

publicNov 2016View details →
geo24/100

Gene expression profile of MDA-MB-231 breast cancer cells with BRCA1 pseudogene (BRCA1P1) knockout (KO)

GEO Series GSE112573. Homo sapiens. 4 samples. Type: Expression profiling by array.

openGEO-OpenMar 2021View details →
geo24/100

Hybrid sequencing characterizes expression and function of mouse pseudogenes

GEO Series GSE176018. Mus musculus. 35 samples. Type: Expression profiling by high throughput sequencing.

openGEO-OpenFeb 2022View details →
geo24/100

Pseudogene repair driven by selection pressure applied in experimental evolution

GEO Series GSE122779. Escherichia coli K-12. 16 samples. Type: Expression profiling by high throughput sequencing.

openGEO-OpenDec 2018View details →
geo20/100

Differentially Expressed Pseudogenes in HIV-1 Infection

GEO Series GSE70785. Homo sapiens. 2 samples. Type: Expression profiling by high throughput sequencing.

openGEO-OpenFeb 2016View details →
geo20/100

Pseudogene INTS6P1 regulates its cognate gene INTS6 through competitive binding of miR-17-5p in hepatocellular carcinoma [miRNA]

GEO Series GSE64632. Homo sapiens. 6 samples. Type: Non-coding RNA profiling by array.

openGEO-OpenJan 2015View details →
geo20/100

Exhaustive profiling in Arabidopsis reveals abundant polysome-associated 24-nt small RNAs including hitherto undescribed AGO5-associated pseudogene-derived siRNAs (psiRNAs) [RNA]

GEO Series GSE99826. Arabidopsis thaliana. 4 samples. Type: Expression profiling by high throughput sequencing.

openGEO-OpenDec 2017View details →
geo20/100

Enrichment of processed pseudogene transcripts in L1-ribonucleoprotein particles

GEO Series GSE43801. Homo sapiens. 5 samples. Type: Expression profiling by high throughput sequencing; Other.

openGEO-OpenMay 2013View details →

ScienceDex guides

Understand access before you commit

These curated guides explain access requirements, typical timelines, costs, and reuse considerations for widely used research datasets.

Compare curated datasets

Allen Brain Atlas

Allen Brain Atlas is an Allen Institute collection of brain map atlases, datasets, APIs, and analysis tools covering mouse, human, and non-human primate brain resources.

allen-brain-atlas
neuroscienceopenDocumentation, web resources, and API references are available online.
Last verified 2026-04-30Open record

Annotated Behaviour and Observability Dataset (ABODe)

ABODe is a University of Edinburgh DataShare dataset for behavior classification in group-housed mice using home-cage video, identities, bounding boxes, ground-plate positions, and annotator labels.

abode-home-cage
behavioral-neuroscienceopenThe DataShare record exposes download links for annotations, documentation, license text, and the zipped per-snippet data directory.
Last verified 2026-04-30Open record

DANDI Archive for NWB datasets

DANDI is a BRAIN Initiative archive for publishing and sharing neurophysiology data, including electrophysiology, optophysiology, and behavioral data packaged as NWB and related standards.

dandi-nwb
electrophysiologyopenPublished Dandiset metadata and archive endpoints are available through the production DANDI API.
Last verified 2026-04-30Open record

International Brain Laboratory public data

The International Brain Laboratory public data releases expose standardized mouse decision-making experiments, including Neuropixels recordings, widefield calcium imaging, behavior, and session metadata accessed through the ONE API.

ibl
behavioral-neuroscienceopenPublic sessions can be searched and loaded from the IBL public data server through ONE.
Last verified 2026-04-29Open record

OpenNeuro

OpenNeuro is a free, open platform for sharing neuroimaging datasets, with public search, dataset pages, and download paths for web, S3, DataLad, and the OpenNeuro CLI.

openneuro
neuroscienceopenPublished datasets are available on demand over the internet.
Last verified 2026-04-29Open record