Find research datasets worth reusing
Search datasets from major research repositories and use ShareScore to quickly assess how well each record supports discovery, access, and reuse.
233
datasets available to search
ShareScore release 0.9.0
Dataset results
233 results for “long-read”
Long-read methylation data analysis with NanoMethViz and Bioconductor
<p>Example data for Bioc2024 workshop. Contains ModBam files and sample annotations. Data comes from mouse neural stem cells aligned to mm10. BAM files have been split into haplotypes and subset to chromosome 7.</p>
Seq2scFv: a toolkit for the comprehensive analysis of display libraries from long-read sequencing platforms
<p>Annotation of full-length scFvs sequenced with PacBio using Seq2scFv (https://github.com/ngs-ai-org/seq2scfv). </p> <p>Raw files were obtained from publicly available dataset generated by Nannini and colleagues in "Combining phage display with SMRTbell next-generation sequencing for the rapid discovery of functional scFv fragments." <em>MAbs</em>. Vol. 13. No. 1. Taylor & Francis, 2021.</p> <p>The results of the annotation are organized in directories. These correspond to the different steps of the tutorial presented in the Seq2scFv repository. </p> <p>├── 1.Preprocessed<br>├── 2.Catalogued<br>├── 3.vscan<br>├── 4.scFv_split<br>├── 5.Linkers<br>├── 6.Flags<br>├── 7.Counts<br>└── focus_libs.txt</p> <p>This entry also contains the code provided in the GitHub repository.</p>
Utilizing Long-read Sequencing to Investigate the EGFR Landscape of EGFR Positive Lung Cancer Patients
ClinicalTrials.gov study NCT06659458. IPD Sharing: YES. Countries: 1. Publications: 0.
Profiling epigenetic aging at cell-type resolution through long-read sequencing
GEO Series GSE296282. Homo sapiens. 5 samples. Type: Methylation profiling by high throughput sequencing.
Nanopore long-reads sequencing of hepatitis A virus RNAs
GEO Series GSE293395. Homo sapiens. 6 samples. Type: Expression profiling by high throughput sequencing.
Direct long-read RNA sequencing of Raji cells
GEO Series GSE243920. Homo sapiens. 1 samples. Type: Expression profiling by high throughput sequencing.
High-resolution transcriptome analysis with long-read RNA sequencing
GEO Series GSE57862. Homo sapiens. 2 samples. Type: Expression profiling by high throughput sequencing.
RNA-Seq analysis (long-reads) of human parental M238P and BRAF inhibitors-resistant M238R melanoma cells
GEO Series GSE171882. Homo sapiens. 2 samples. Type: Non-coding RNA profiling by high throughput sequencing.
Nanopore long-read RNA sequencing in untreated (UN), sodium arsenite (SA) and heat shock (HS) stressed HeLa cells
GEO Series GSE277764. Homo sapiens. 12 samples. Type: Expression profiling by high throughput sequencing; Other.
Accurate isoform quantification by joint short- and long-read RNA-sequencing [long reads]
GEO Series GSE271527. Homo sapiens. 6 samples. Type: Expression profiling by high throughput sequencing.
Long-read single-cell RNA transcriptomics reveals a signature of PIK3CA mutation in a human capillary malformation
GEO Series GSE253153. Homo sapiens. 1 samples. Type: Expression profiling by high throughput sequencing.
Optimizing Single-Cell Long-Read Sequencing for Enhanced Isoform Detection in Pancreatic Islets [sc_islet]
GEO Series GSE295353. Mus musculus. 7 samples. Type: Expression profiling by high throughput sequencing.
Long-read transcriptome profiling of human lung cancer cell lines
GEO Series GSE227000. Homo sapiens. 5 samples. Type: Expression profiling by high throughput sequencing.
Optimizing Single-Cell Long-Read Sequencing for Enhanced Isoform Detection in Pancreatic Islets [bulk_islet]
GEO Series GSE295351. Mus musculus. 4 samples. Type: Expression profiling by high throughput sequencing.
Nanopore long-read sequencing from Nutlin-3a-treated MCF-7 cells
GEO Series GSE226080. Homo sapiens. 2 samples. Type: Expression profiling by high throughput sequencing.
Integrative genotyping of cancer and immune phenotypes by long-read sequencing
GEO Series GSE243227. Homo sapiens. 13 samples. Type: Expression profiling by high throughput sequencing.
Long-read RNA sequencing of cattle, pig, and chicken tissues
GEO Series GSE160028. Gallus gallus; Bos taurus; Sus scrofa. 190 samples. Type: Expression profiling by high throughput sequencing.
Long-read transcriptome sequencing reveals expression characteristics of osteosarcoma
GEO Series GSE218035. Homo sapiens. 36 samples. Type: Expression profiling by high throughput sequencing.
Cataloging the potential functional diversity of Cacna1e splice variants using long-read sequencing
GEO Series GSE295903. Rattus norvegicus. 4 samples. Type: Expression profiling by high throughput sequencing.
Generation of full-length circRNA libraries for Oxford Nanopore long-read sequencing
GEO Series GSE197872. Homo sapiens. 4 samples. Type: Expression profiling by high throughput sequencing.
ScienceDex guides
Understand access before you commit
These curated guides explain access requirements, typical timelines, costs, and reuse considerations for widely used research datasets.
Allen Brain Atlas
Allen Brain Atlas is an Allen Institute collection of brain map atlases, datasets, APIs, and analysis tools covering mouse, human, and non-human primate brain resources.
Annotated Behaviour and Observability Dataset (ABODe)
ABODe is a University of Edinburgh DataShare dataset for behavior classification in group-housed mice using home-cage video, identities, bounding boxes, ground-plate positions, and annotator labels.
DANDI Archive for NWB datasets
DANDI is a BRAIN Initiative archive for publishing and sharing neurophysiology data, including electrophysiology, optophysiology, and behavioral data packaged as NWB and related standards.
International Brain Laboratory public data
The International Brain Laboratory public data releases expose standardized mouse decision-making experiments, including Neuropixels recordings, widefield calcium imaging, behavior, and session metadata accessed through the ONE API.
OpenNeuro
OpenNeuro is a free, open platform for sharing neuroimaging datasets, with public search, dataset pages, and download paths for web, S3, DataLad, and the OpenNeuro CLI.