Skip to main content
Powered by ShareScore

Find research datasets worth reusing

Search datasets from major research repositories and use ShareScore to quickly assess how well each record supports discovery, access, and reuse.

1,063

datasets available to search

ShareScore release 0.9.0

Reset

Dataset results

1,063 results for “Search”

Learn how ShareScore rates datasets ↗
zenodo36/100

ANION GAP OR SERUM LACTATE-IN SEARCH OF A BETTER PROGNOSTIC MARKER IN SEPSIS A CROSS-SECTIONAL STUDY IN A RURAL TERTIARY CARE HOSPITAL

<p>master data sheet</p>

opencc-by-4.0Apr 2023View details →
zenodo36/100

Persistent and occasional: searching for the variable population of the ZTF/4MOST sky using ZTF data release 11.

<p>In this dataset we provide classifications of Zwicky Transient&nbsp;Facility (ZTF) Data Release 11 (DR11) light curves, from the work &quot;Persistent and occasional: searching for the variable population of the&nbsp;ZTF/4MOST sky using ZTF data release 11&quot;, accepted for publication in the Astronomy and Astrophysics Journal (S&aacute;nchez-S&aacute;ez et al. 2023). Here we provide&nbsp;classifications for objects in the ZTF/4MOST sky, including&nbsp;86,576,577 sources in the g band and 140,409,824 in&nbsp;the r band. The classifications are provided in&nbsp;parquet files, separated by class and ZTF band. We also provide the labeled sets used to train the models for&nbsp;each band, and the master catalog used to construct the labeled sets.</p> <p>&nbsp;</p> <p>File description:</p> <p>Classifications: files with names&nbsp;%class%_cand_%band%.parquet.gz</p> <p>Labeled set g band:&nbsp;LS_ZTFg.parquet.gz</p> <p>Labeled set r&nbsp;band:&nbsp;LS_ZTFr.parquet.gz</p> <p>Master catalog:&nbsp;mast_cat.parquet.gz</p>

opencc-by-4.0Apr 2023View details →
zenodo36/100

TIRF Microscopy Data Files. "Multiple RNA- and DNA-binding proteins exhibit direct transfer of polynucleotides: Implications for target site search"

<p>TIRF-microscopy images and analysis data from single-molecule experiments assessing the direct transfer phenomenon in the TREX1 exonuclease.&nbsp;</p>

opencc-by-4.0Dec 2021View details →
zenodo36/100

Artifact supplement for 'Search and Explore: Symbiotic Policy Synthesis in POMDPs'

<p>Artifact supplement for &#39;Search and Explore: Symbiotic Policy Synthesis in POMDPs&#39;&nbsp; (CAV 2023).<br> <br> This repository includes:<br> -&nbsp;the&nbsp;Docker image containing the software used in the experiments along with the investigated benchmarks and original log files<br> - a guide for replicating the experiments presented in the publication<br> - two versions of the paper: the initial submission and the final version with the revised experiments</p>

opencc-by-4.0Apr 2023View details →
zenodo36/100

Supplementary Data: Simple, but not simplified: A new approach for optimising beyond-Standard Model physics searches at the Large Hadron Collider

<p><strong>Supplementary Data</strong></p> <p><em>Simple, but not simplified: A new approach for optimising beyond-Standard Model physics searches at the Large Hadron Collider</em></p> <p>This record contains the full dataset generated for the study &quot;<em>Simple, but not simplified: A new approach for optimising beyond-Standard Model physics searches at the Large Hadron Collider</em>&quot;. It contains the following files:</p> <ul> <li>data_cross_sections_full_exact.csv - a CSV file with the soft breaking parameters M1, M2, mu and tanb, the neutralino/chargino masses, the cross sections for neutralino-neutralino, chargino-neutralino and chargino-chargino production at the LHC operating at a CM energy of 13 TeV, and the branching ratios of the unstable charginos/neutralinos.&nbsp;</li> <li>benchmark_points.zip - this zip file contains the SLHA files for the four benchmark points as shown in the paper, together with the prospino output for the cross section.&nbsp;</li> </ul> <p>&nbsp;</p>

opencc-by-4.0May 2023View details →
zenodo36/100

Exploratory Search Workflows (ESW) collection

<p>This Zenodo digital object represents the dataset of the&nbsp;Exploratory Search Worklflows (ESW) collection. It contains the ontology (owl file) and a folder with the exploratory workflows divided by &quot;track&quot;.&nbsp;Each track&nbsp;contains the query logs, the exploratory workflows execution,&nbsp;the exploratory workflows evaluated, the ground truths and the serialized turtle files.</p>

opencc-by-4.0May 2023View details →
zenodo36/100

MSFragger open searches of proteome shotgun and phosphoproteomic runs PXD013868 (Mergner et al., 2020)

<p>MSFragger open searches of phosphoproteomics and shotgun proteome runs of large-scale tissue atlas in Arabidopis (Mergner et al., 2020). Part of the Plant PTM Viewer 2.0 update paper.</p>

opencc-by-4.0Jun 2023View details →
zenodo36/100

Literature search datasets

<p>The datasets contain key results of literature search for the review paper to which this data set is linked.</p>

opencc-by-4.0Oct 2022View details →
zenodo36/100

In Search of Caribbeanness: Explorations of the Skin Ego in David Boxer and Stan Musquer's Works

<p>In this conference paper,&nbsp;Fr&eacute;d&eacute;ric LEFRAN&Ccedil;OIS, Doctor of Literature, discusses the concept of the skin-ego. The human epidermis, the outer surface of the soul, has become the seat of conflict between Europe, Africa and Asia. To develop the concept, he sets out to answer a question: &quot;Is the skin-self a protective or alienating envelope? &quot;His answer is based on the work of David BOXER and Stan MUSQUER. In his view, the skin-self is an agent of inclusion or exclusion, depending on the socio-cultural or ethnic context in which an individual finds himself.</p>

opencc-by-4.0Jul 2023View details →
zenodo36/100

Petascale Homology Search for Structure Prediction - MSAs

<p>Multiple sequence alignments (MSAs) of CASP15 targets used in &quot;Petascale Homology Search for Structure Prediction&quot; publication.&nbsp;</p> <p>Each MSA is constructed from a combination of different sequence databases, using different search tools, 1) ColabFoldDB (CFDB) using ColabFold search module, 2) Sequence Read Archive (SRA) using MMseqs2, and 3) UniRef30+BFD using HHblits (HH).</p> <ul> <li><strong>Queries</strong>: Regular CASP15 TS targets, excluding TBM-easy</li> <li><strong>MSAs (A3M)&nbsp;</strong> <ul> <li>CFDB (cfdb.tar.gz)</li> <li>SRA + CFDB (sra_cfdb.tar.gz)</li> <li>HH + CFDB (hh_cfdb.tar.gz)</li> <li>HH + SRA + CFDB (hh_sra_cfdb.tar.gz)</li> </ul> </li> </ul>

opencc-by-4.0Jul 2023View details →
zenodo36/100

DeepCV: A Deep Learning Framework for Blind Search of Collective Variables in Expanded Configurational Space

<p>We present <em>Deep learning for Collective Variables</em> (DeepCV), a computer code that provides an efficient and customizable implementation of the deep autoencoder neural network (DAENN) algorithm that has been developed in our group for computing collective variables (CVs) and can be used with enhanced sampling methods to reconstruct free energy surfaces of chemical reactions. DeepCV can be used to conveniently calculate molecular features, train models, generate CVs, validate rare events from sampling, and analyze a trajectory for chemical reactions of interest. We use DeepCV in an example study of the conformational transition of cyclohexene, where metadynamics simulations are performed using DAENN-generated CVs. The results show that the adopted CVs give free energies in line with those obtained by previously developed CVs and experimental results. DeepCV is open-source software written in Python/C++ object-oriented languages, based on the TensorFlow framework and distributed free of charge for noncommercial purposes, which can be incorporated into general molecular dynamics software. DeepCV also comes with several additional tools, i.e., an application program interface (API), documentation, and tutorials.</p>

opencc-by-4.0Nov 2022View details →
zenodo36/100

Search strategies for generic justification of 18F-PSMA PET/CT in the staging of primary prostate cancer and the restaging of recurrent and metastatic prostate cancer

<p>The dataset includes the complete, reproducible search strategies for all literature databases searched during this project. The search strategies address the following research questions:</p> <p>RQ 1: In the primary staging of patients with high risk prostate cancer, what is the sensitivity, specificity and diagnostic accuracy of 18F-PSMA?</p> <p>RQ 2:&nbsp;In patients with biochemically recurrrent prostate cancer, what is the sensitivity, specificity and diagnostic accuracy of 18F-PSMA?</p> <p>RQ 3:&nbsp;What is the risk of adverse events (including dose) associated with receiving 18F-PSMA?</p> <p>RQ 4: Is 18F-PSMA uptake associated with response to treatments for metastatic prostate cancer?</p>

opencc-by-4.0Jul 2023View details →
zenodo36/100

Venom: Toxin Accessory Domain Search

<p>This contains the different fasta files used and generated for the search of accessory domains fused with toxin proteins in a variety of venoms and tick saliva.&nbsp;</p> <p>2023-06-15-updated-venom-tick-proteins.fasta :&nbsp; fasta file of our custom &quot;venom proteins from transcriptomes&rdquo; data set, that was generated from&nbsp;the transcriptomes of&nbsp;venom glands of 124 species and from the salivary glands of 21 tick species</p> <p>uniprot-toxin-clustering_rep_seq.fasta: fasta file of all the representative sequences of the UniProt manually curated database of proteins and toxins from various venoms (from the&nbsp;animal toxin annotation project ) obtained after clustering.</p> <p>2023-06-20-outliers-accessory-hits.fasta: fasta file of the accessory sequences identified on the toxin outliers observed in venoms and saliva.</p> <p>updated-accessory-toxin-clustering_rep_seq.fasta: fasta file of the representative sequences of the toxin outliers accessory sequences obtained after clustering.</p> <p>&nbsp;</p>

opencc-by-4.0Aug 2023View details →
zenodo36/100

Search strategies for herpes zoster vaccination for adults

<p>The dataset includes the complete, reproducible search strategies for all literature databases searched during this project. The search strategies were designed to answer the following research questions:</p> <ul> <li>RQ 1 &ndash; What is the clinical efficacy and effectiveness of the currently licensed and approved recombinant vaccine for the prevention of herpes zoster and associated complications, in adults aged 50 years and older and in adults aged 18 and over who are at increased risk of herpes zoster?</li> <li>RQ 2 &ndash; What is the safety profile of the currently licensed and approved recombinant vaccine for the prevention of herpes zoster in adults aged 50 years and older and in adults aged 18 and over who are at increased risk of herpes zoster?</li> </ul> <p>&nbsp;</p>

opencc-by-4.0Aug 2023View details →
zenodo36/100

Search strategies for 177 lu-PSMA in the treatment of metastatic castrate-resistant prostate cancer

<p>The dataset includes the complete, reproducible search strategies for all literature databases searched during this project. The search strategies were designed to answer the following research questions:</p> <p><strong>RQ1.</strong> Does the use of radioligand therapy using <sup>177</sup>Lu-PSMA lead to improved overall survival and progression-free survival, compared with other available treatment(s) in patients with metastatic, castrate-resistant prostate cancer?</p> <p><strong>RQ2.</strong> Does the use of radioligand therapy using <sup>177</sup>Lu-PSMA lead to improved quality of life or symptom control, compared with other available treatment(s), in patients with metastatic, castrate-resistant prostate cancer?</p> <p><strong>RQ3. </strong>&nbsp;What is the risk of adverse events and toxicity associated with radioligand therapy using <sup>177</sup>Lu-PSMA, compared with other available treatment&nbsp;in patients with metastatic, castrate-resistant prostate cancer?</p>

opencc-by-4.0Aug 2023View details →
zenodo36/100

A Novel Algorithm for Estimating Web Page Ranking in Search Engine Results Pages

<p><em><strong>Abstract:</strong> </em>Search engine optimization (SEO) can make a big improvement in the traffic to a web page. Because search engines keep their main rules of ranking undeclared, it&rsquo;s important to develop models that can estimate the ranking of a web page in the search engine to be able to optimize web pages to rank higher in the search engine. The available research methodologies used machine learning algorithms to provide solutions for this target with the help of generated datasets by scraping the search engine results pages (SERP) and crawling web pages. Their proposed models suffered from the inability to be updated dynamically if the search engine updated its ranking algorithm, and their input data did not include the diversity of web pages and languages. This research will propose a novel original rank estimation algorithm that&rsquo;s able to overcome other research challenges, with a set of comparative experiments and complexity analysis. Results will show that the proposed algorithm could achieve higher values of accuracy, precision, and recall.</p> <p><strong><em>Dataset:&nbsp;</em></strong></p> <p>For research purpose, the dataset will play two roles, first, it will act the role of search engine result pages (SERP), and second, it will be used to test algorithms and calculate performance measurements.&nbsp;Dataset is consisting of 9930 web pages, aimed to identify search results pages, focusing on the top 3 pages of SERP, with 31 extracted attributes that&#39;s related to search engine optimization (SEO). The distribution of examples between class labels was balanced, with changes due to scraping operation issues, but not significantly different, with fractions of 39.9%, 34.6%, and 25.5% for the class labels page1, page2, and page 3. Feature names are: &#39;Title 1 Length&#39;, &#39;Title 2 Length&#39;, &#39;Meta Description 1 Length&#39;, &#39;Meta Description 2 Length&#39;, &#39;Meta Keywords 1 Length&#39;, &#39;H1-1 Length&#39;, &#39;H1-2 Length&#39;, &#39;H2-1 Length&#39;, &#39;H2-2 Length&#39;, &#39;Size (bytes)&#39;, &#39;Word Count&#39;, &#39;Text Ratio&#39;, &#39;Inlinks&#39;, &#39;Unique Inlinks&#39;, &#39;Unique JS Inlinks&#39;, &#39;% of Total&#39;, &#39;Outlinks&#39;, &#39;Unique Outlinks&#39;, &#39;Unique JS Outlinks&#39;, &#39;External Outlinks&#39;, &#39;Unique External Outlinks&#39;, &#39;Unique External JS Outlinks&#39;, &#39;Response Time&#39;, &#39;Status Code&#39;, &#39;Keyword in MetaDescription1&#39;, &#39;Keyword in Title1&#39;, &#39;Keyword in MetaKeywords1&#39;, &#39;Keyword in URL&#39;, &#39;Has LastModified&#39;, &#39;Keyword in Headers&#39;, and &#39;Keyword in Emphasized Text&#39;.</p> <p>The process of dataset generation involved&nbsp;scraping the search engine, extracting URLs for selected keywords, focusing on feature extraction, cleaning and preprocessing, and generating new attributes related to keywords in web pages. It&nbsp;involved also removing missing values, duplicates, and data type conversions to obtain a comprehensive dataset.<br> Keyword selection involves selecting keywords from various categories and considering diversity, including high and low traffic, long-term and short-term keywords, and generic and branded keywords. Apify online tool was used for search engine scraping with default language and US country, resulting in 388 selected keywords with 30 results per keyword. Dataset included extracted SEO features from 9991 web pages using screamingFrog desktop software and Rapidminer desktop software, determining page SEO-friendliness and comparing it to SERP rankings. Dataset cleaning involved removing redundant attributes, removing paid SERP results, replacing missing values, and converting data types. Rapidminer was used for data cleaning and preprocessing, generating new attributes related to keyword usage in web pages.<br> &nbsp;</p>

opencc-by-4.0Sep 2023View details →
ClinicalTrials.gov36/100

Conduct Disorder in Prison: Search for Prospective and Retrospective Recurring and Contextual Elements

ClinicalTrials.gov study NCT02493088. IPD Sharing: Not stated. Countries: 1. Publications: 4.

restrictedIPD-UNDECIDEDFeb 2026View details →
ClinicalTrials.gov36/100

Melanoma of the Skin and Exposure to Solar Ultraviolet Radiation at Work in Modena Territory: a Case-control Study to Promote an Active Search and Prevention of Occupational Diseases Based on Recent I

ClinicalTrials.gov study NCT07251335. IPD Sharing: NO. Countries: 1. Publications: 7.

closedIPD-NOFeb 2026View details →
ClinicalTrials.gov36/100

SEARCH SAPPHIRE Phase A: A Multisectoral Strategy to Address Persistent Drivers of the HIV Epidemic in East Africa

ClinicalTrials.gov study NCT04810650. IPD Sharing: NO. Countries: 2. Publications: 7.

closedIPD-NOFeb 2026View details →
ClinicalTrials.gov36/100

SEARCH CAB LA Dynamic Choice HIV Prevention Study Extension

ClinicalTrials.gov study NCT05549726. IPD Sharing: NO. Countries: 2. Publications: 1.

closedIPD-NOFeb 2026View details →

ScienceDex guides

Understand access before you commit

These curated guides explain access requirements, typical timelines, costs, and reuse considerations for widely used research datasets.

Compare curated datasets

Allen Brain Atlas

Allen Brain Atlas is an Allen Institute collection of brain map atlases, datasets, APIs, and analysis tools covering mouse, human, and non-human primate brain resources.

allen-brain-atlas
neuroscienceopenDocumentation, web resources, and API references are available online.
Last verified 2026-04-30Open record

Annotated Behaviour and Observability Dataset (ABODe)

ABODe is a University of Edinburgh DataShare dataset for behavior classification in group-housed mice using home-cage video, identities, bounding boxes, ground-plate positions, and annotator labels.

abode-home-cage
behavioral-neuroscienceopenThe DataShare record exposes download links for annotations, documentation, license text, and the zipped per-snippet data directory.
Last verified 2026-04-30Open record

DANDI Archive for NWB datasets

DANDI is a BRAIN Initiative archive for publishing and sharing neurophysiology data, including electrophysiology, optophysiology, and behavioral data packaged as NWB and related standards.

dandi-nwb
electrophysiologyopenPublished Dandiset metadata and archive endpoints are available through the production DANDI API.
Last verified 2026-04-30Open record

International Brain Laboratory public data

The International Brain Laboratory public data releases expose standardized mouse decision-making experiments, including Neuropixels recordings, widefield calcium imaging, behavior, and session metadata accessed through the ONE API.

ibl
behavioral-neuroscienceopenPublic sessions can be searched and loaded from the IBL public data server through ONE.
Last verified 2026-04-29Open record

OpenNeuro

OpenNeuro is a free, open platform for sharing neuroimaging datasets, with public search, dataset pages, and download paths for web, S3, DataLad, and the OpenNeuro CLI.

openneuro
neuroscienceopenPublished datasets are available on demand over the internet.
Last verified 2026-04-29Open record