Find research datasets worth reusing
Search datasets from major research repositories and use ShareScore to quickly assess how well each record supports discovery, access, and reuse.
302
datasets available to search
ShareScore release 0.9.0
Dataset results
302 results for “protein sequence”
Fig. 3 in Protein sequences from mastodon and Tyrannosaurus rex revealed by mass spectrometry
Fig. 3. The LC/MS/MS fragmentation pattern from a 68-million-year-old T. rex peptide. (A) The experimental MS/MS spectrum for the T. rex doubly charged hydroxylated tryptic peptide sequence GVQPP(OH)GPQGPR from femur bone extract identified by LC/MS/MS. (B) The synthetic version of the same sequence. All major fragment ions from the experimental spectrum are in very good alignment with ions from the synthetic version, confirming the sequence. This molecular sequencing evidence of protein from a 68-million-year-old fossilized bone demonstrates excellent preservation of the T. rex femur and the high sensitivity of state-of-the-art MS technology.
Fig. 1 in Comment on "Protein Sequences from Mastodon and Tyrannosaurus rex Revealed by Mass Spectrometry"
Fig. 1. Plot of radiocarbon age versus estimated effective collagen degradation temperature for radiocarbon-dated bones from laboratory databases (principally Oxford and Groningen). The line represents the expected calendar age at which 1% of the original collagen remains following a zero-order reaction; almost no bone collagen survives beyond this predicted limit. (Inset) The 99% confidence intervals of amino acid compositions by first two principal component analyses (48% of total variance) for bones from NW Europe aged <11 ky (n = 324), 11 to 110 ky (n = 210), 110 to 130 ky (n = 26), and 130 to 700 ky (n = 31). Pliocene samples are not plotted, as their composition (n = 8) is highly variable and yields of amino acids are low. The orange line indicates a compositional trend observed when compact bone is heated for 32 days at 95°C, which reduces collagen to 1% of the initial concentration [each inflection represents a separate analysis; n = 32)]. The composition becomes more similar to mixed tissue samples (meat and bone meal; n = 32), principally due to the depletion of Gly. An amino acid profile for mammoth is consistent with collagen, unlike the associated sediment sample [data from (11)].
Evolutionary adaptation of the chromodomain of the HP1 protein Rhino allows th eintegration of heterochromatin and DNA sequence signals
GEO Series GSE244196. Drosophila melanogaster. 17 samples. Type: Genome binding/occupancy profiling by high throughput sequencing; Non-coding RNA profiling by high throughput sequencing.
Transcriptome Sequence Analysis of Pediatric Acute Megakaryoblastic Leukemia Identifies An Inv(16)(p13.3;q24.3)-Encoded CBFA2T3-GLIS2 Fusion Protein As a Recurrent Lesion in 39% of Non-Infant Cases
GEO Series GSE35203. Homo sapiens. 43 samples. Type: Expression profiling by array.
Highly quantitative measurement of differential protein-genome binding with PerCell chromatin sequencing
GEO Series GSE271986. Homo sapiens. 20 samples. Type: Genome binding/occupancy profiling by high throughput sequencing.
Polymerase pausing induced by sequence-specific RNA binding protein drives heterochromatin assembly (NET-Seq)
GEO Series GSE114536. Schizosaccharomyces pombe. 6 samples. Type: Other.
Global profiling of the RNA and protein complexes of Escherichia coli by size exclusion chromatography followed by RNA sequencing and mass spectrometry (SEC-seq)
GEO Series GSE212408. Escherichia coli. 21 samples. Type: Expression profiling by high throughput sequencing.
RNA sequencing of Tau overexpressing human SH-SY5Y cell line revels that Tau protein is able to alter global gene expression and chromatin structure
GEO Series GSE239956. Homo sapiens. 6 samples. Type: Expression profiling by high throughput sequencing.
An engineered decoy receptor for SARS-CoV-2 broadly binds protein S sequence variants
GEO Series GSE159372. Severe acute respiratory syndrome coronavirus 2. 7 samples. Type: Other.
Transcriptome Sequence Analysis of Pediatric Acute Megakaryoblastic Leukemia Identifies An Inv(16)(p13.3;q24.3)-Encoded CBFA2T3-GLIS2 Fusion Protein As a Recurrent Lesion in 39% of Non-Infant Cases [2
GEO Series GSE35201. Homo sapiens. 29 samples. Type: Expression profiling by array.
RNA sequencing (RNA-seq) for identifing differentially expressed genes for mitochondrial unfolded protein response in Arabidopsis
GEO Series GSE198496. Arabidopsis thaliana. 9 samples. Type: Expression profiling by high throughput sequencing.
Deep sequencing after alcelaphine gammaherpesvirus 1 infection reveals the nature of CD8+ T cell expansion and identify an essential viral protein for fatal bovine malignant catarrhal fever
GEO Series GSE253729. Bos taurus. 31 samples. Type: Expression profiling by high throughput sequencing; Genome binding/occupancy profiling by high throughput sequencing.
Transcriptomic analysis of two 14-3-3 proteins Bmh1 and Bmh2 under osmotic and oxidative stresses in the entomopathogen fungus Beauveria bassiana by using RNA sequencing
GEO Series GSE58688. Beauveria bassiana. 6 samples. Type: Expression profiling by high throughput sequencing.
Long-Read Sequencing and Proteomics Reveal Blood Transcriptome and Protein Expression Profiles in Multiple Primary Lung Cancers
GEO Series GSE311391. Homo sapiens. 12 samples. Type: Expression profiling by high throughput sequencing.
Isolation of biologically active small peptide from KSHV LANA protein sequence, which induces CHD4 degradation to promote cell differentiation and inhibits leukemic cell growth
GEO Series GSE228657. Homo sapiens. 30 samples. Type: Expression profiling by high throughput sequencing; Other.
DNA-protein Crosslinking Sequencing for Genome-wide Mapping of Thymidine Glycol
GEO Series GSE184204. Homo sapiens. 6 samples. Type: Other.
Structural stability based deep sequencing of 5 protein targets
GEO Series GSE248664. Escherichia coli. 28 samples. Type: Other.
Deep sequencing after alcelaphine gammaherpesvirus 1 infection reveals the nature of CD8+ T cell expansion and identify an essential viral protein for fatal bovine malignant catarrhal fever [RNA-seq]
GEO Series GSE253728. Bos taurus. 14 samples. Type: Expression profiling by high throughput sequencing.
Targeted DNA and surface protein sequencing at single cell level of BM MNCs and PB MNCs from CCUS patients
GEO Series GSE276492. Homo sapiens. 5 samples. Type: Other.
Gene and protein sequence features augment HLA class I ligand predictions (ribosome profiling)
GEO Series GSE210998. Homo sapiens. 6 samples. Type: Other.
ScienceDex guides
Understand access before you commit
These curated guides explain access requirements, typical timelines, costs, and reuse considerations for widely used research datasets.
Allen Brain Atlas
Allen Brain Atlas is an Allen Institute collection of brain map atlases, datasets, APIs, and analysis tools covering mouse, human, and non-human primate brain resources.
Annotated Behaviour and Observability Dataset (ABODe)
ABODe is a University of Edinburgh DataShare dataset for behavior classification in group-housed mice using home-cage video, identities, bounding boxes, ground-plate positions, and annotator labels.
DANDI Archive for NWB datasets
DANDI is a BRAIN Initiative archive for publishing and sharing neurophysiology data, including electrophysiology, optophysiology, and behavioral data packaged as NWB and related standards.
International Brain Laboratory public data
The International Brain Laboratory public data releases expose standardized mouse decision-making experiments, including Neuropixels recordings, widefield calcium imaging, behavior, and session metadata accessed through the ONE API.
OpenNeuro
OpenNeuro is a free, open platform for sharing neuroimaging datasets, with public search, dataset pages, and download paths for web, S3, DataLad, and the OpenNeuro CLI.