Skip to main content
Powered by ShareScore

Find research datasets worth reusing

Search datasets from major research repositories and use ShareScore to quickly assess how well each record supports discovery, access, and reuse.

251

datasets available to search

ShareScore release 0.7.1

Reset

Dataset results

251 results for “deep learning models”

Learn how ShareScore rates datasets ↗
zenodo40/100

Replication Package for "PyTraceBERT: Python Traceback-based Language Model for Detecting Compatibility Issues in Deep Learning Systems"

<p>This package contains the traceback data, pre-trained models, and static word embeddings used in the paper, PyTraceBERT: Python Traceback-based Language Model for Detecting Compatibility Issues in Deep Learning Systems.</p>

opencc-by-4.0Sep 2024View details →
zenodo40/100

Winter Precipitation-Type Models for "Evidential Deep Learning: Enhancing Predictive Uncertainty Estimation for Earth System Science Applications"

<p>This contains trained model weights, scalers, and evaluation metrics for the winter precipitation-type models trained as part of the paper "Evidential Deep Learning: Enhancing Predictive Uncertainty Estimation for Earth System Science Applications".&nbsp;</p>

opencc-by-4.0Sep 2024View details →
zenodo40/100

Enhancing Smartphone Battery Life: A Deep Learning Model Based on User-Specific Application and Network Behaviour

<p>This work presents an analysis based on training AI models directly on devices to make personalized predictions tailored to individual usage patterns, ensuring that each user benefits from a personalized approach to battery management. By integrating these AI-based insights, mobile devices can proactively manage power consumption, improving battery performance and user satisfaction. This personalized, intelligent approach to battery management represents a significant advance in optimizing device efficiency and addresses the growing demand for longer-lasting mobile technology.</p>

opencc-by-4.0Oct 2024View details →
zenodo40/100

Deep learning model for characterizing protein-RNA interactions from sequence at single-base resolution

<p>&nbsp;</p> <p><a href="https://zenodo.org/api/records/14021440/draft/files/encode_eclip.h5/content" target="_blank" rel="noopener noreferrer">encode_eclip.h5</a> - This file contains the training, validation, and test data for the Reformer model.</p> <p><a href="https://zenodo.org/api/records/14021440/draft/files/encode_eclip_bc.h5/content" target="_blank" rel="noopener noreferrer">encode_eclip_bc.h5</a> - This file contains the training, validation, and test data for the Reformer-BC model.</p> <p><a href="https://zenodo.org/api/records/14027315/draft/files/Reformer-code.zip/content" target="_blank" rel="noopener">Reformer-code.zip</a> - This file contains the training code of Reformer.</p>

opencc-by-4.0Oct 2024View details →
zenodo40/100

Data for "Deep learning-based model for diagnosing Alzheimer's disease and tauopathies"

<p>Image datasets and tuned models used in the paper (Koga et al., 2021). Data.zip contains image and text files for training models. Test.zip contains 12 images from 4 patients, which are a part of the hold-out dataset images used in the paper. There are 9 CSV files, which contain the results of tau burden quantification. Python code is available at GitHub (<a href="https://github.com/Koga-MD/DL-Tauopathies">https://github.com/Koga-MD/DL-Tauopathies</a>).&nbsp;</p>

opencc-by-4.0Jul 2021View details →
zenodo40/100

CNN models and training, validation and test datasets for "PlotMI: interpretation of pairwise interactions and positional preferences learned by a deep learning model from sequence data"

<p>Convolutional neural network (CNN) models and their respective training, validation and test datasets used in manuscript:</p> <p>Tuomo Hartonen, Teemu Kivioja and Jussi Taipale, &quot;PlotMI: interpretation of pairwise interactions and positional preferences learned by a deep learning model from sequence data&quot;</p>

opencc-by-4.0Mar 2021View details →
zenodo40/100

Training Deep Learning Models to Estimate SWAT Parameters using Streamflow Observations

<p>This folder provides the&nbsp;simulation and observational data</p> <p>Simulation data using SWAT (1000 realz)<br> Train, Val, and Test splits (80/10/10)<br> Observational data for ARW (WY2000-2016)</p>

opencc-by-4.0Oct 2022View details →
zenodo40/100

Datasets used to train the models in "Deep learning for denoising High-Rate Global Navigation Satellite System data."

<p>Datasets used to train the models in &quot;Deep learning for denoising High-Rate Global Navigation Satellite System data.&quot;&nbsp; Additional information can be found at&nbsp;https://github.com/amtseismo/hrgnss_denoising.</p>

opencc-by-4.0Sep 2022View details →
dryad40/100

Deep learning models challenge the prevailing assumption that face-like effects for objects of expertise support domain-general mechanisms

<p>The question of whether perceptual expertise is mediated by general-expert or domain-specific processing mechanisms has been debated for decades. Because humans are experts in face recognition, face-like neural and cognitive effects for objects of expertise were considered to support for the general-expertise hypothesis. Conversely, stronger effects for faces than objects of expertise were considered to support the domain-specific hypothesis. However, the effects of domain, experience, and level of categorization, are confounded in human studies, which may lead to erroneous inferences. To overcome these limitations, we used computational models of perceptual expertise and tested different domains (objects, faces, birds) and levels of categorization (basic, sub-ordinate, individual) in isolation, matched for amount of experience. Like humans, the models generated a larger inversion effect for faces than for objects. Importantly, a face-like inversion effect was found for individual-based categorization of non-faces (birds) but only in a network specialized for that domain. Thus, contrary to prevalent assumptions, face-like effects in objects of expertise may originate from domain-specific rather than domain-general processing mechanisms. More generally, we show how deep learning algorithms can be used to isolate the effects of factors that are inherently confounded in the natural environment of biological organisms.</p>

opencc-zeroApr 2023View details →
zenodo40/100

Training patches and prediction codes of deep learning (LANA) model for Landsat 8/9 cloud/shadow mask

<p>This dataset includes (i) the image patches dataset and (ii) application/prediction (not training) codes for Landsat 8 cloud and cloud shadow masking used in a paper in review and uploaded here: &nbsp;</p> <p>Hankui Zhang, Dong Luo, David Roy, A learning attention network algorithm (LANA) for accurate Landsat-8 cloud and shadow masking,&nbsp;<em>Remote Sensing of Environment</em>&nbsp;</p> <p>The documentation is in&nbsp;<a href="https://zenodo.org/api/files/5462baa5-2bba-4b0f-92aa-c17681b6464b/l8_training_data_readme_new.pdf?versionId=95253efb-6447-46d8-ae4b-ac0c93b43532">l8_training_data_readme_new.pdf</a>.&nbsp;</p>

opencc-by-4.0Apr 2023View details →
zenodo40/100

Image dataset for cow identification, including code to train deep learning model, as well as analysis of results (SmARtview, 51088)

<p>This dataset and code was a result of the UKRI project &quot;SmARtview: An AI-powered Augmented Reality Tool for Animal Health and Productivity&quot;, linked here:&nbsp;<a href="https://gtr.ukri.org/projects?ref=51088">https://gtr.ukri.org/projects?ref=51088</a></p> <p>These files are intended to be used for an accompanying publication in an academic journal.</p> <p>Anyone is free to use the contents for research and teaching&nbsp;purposes.</p>

opencc-by-4.0May 2023View details →
dryad40/100

Deep learning models challenge the prevailing assumption that face-like effects for objects of expertise support domain-general mechanisms

Open the record for dataset details and reuse information.

publicApr 2023View details →
dryad40/100

phyddle: Software for exploring phylogenetic models with deep learning

Open the record for dataset details and reuse information.

publicSep 2025View details →
dryad40/100

Data and code from: Learning a deep language model for microbiomes: The power of large scale unlabeled microbiome data

Open the record for dataset details and reuse information.

publicFeb 2025View details →
zenodo36/100

Benign samples used in article "DeepDetectNet vs RLAttackNet: An Adversarial Method to Improve Deep Learning-based Static Malware Detection Model"

<p>This repository contains all benign samples used in article &quot;DeepDetectNet vs RLAttackNet: An Adversarial Method to Improve Deep Learning-based Static Malware Detection Model&quot;. It is safe to download these samples.</p>

opencc-by-4.0Feb 2020View details →
zenodo36/100

Anonymized Dataset for "Towards a Better Understanding of Reverse-Complement Equivariance for Deep Learning Models in Genomics"

<p>Anonymous dataset for the paper&nbsp;&quot;Towards a Better Understanding of Reverse-Complement Equivariance for Deep Learning Models in Genomics.&quot; Includes data for simulated, binary prediction, and profile prediction tasks.&nbsp;</p>

opencc-by-4.0Jan 2021View details →
zenodo36/100

Data and trained word2vec model for ``Easy over Hard: A Case Study on Deep Learning''

<p>The data include: training  and testing data pairs</p> <p>The  word2vec model is pre-trained. </p> <p>More details, please refer to the paper</p>

opencc-by-4.0Mar 2017View details →
zenodo36/100

Pre-trained word2vec models for ``Easy over Hard: A Case Study on Deep Learning''

<p>Since the whole stack overflow dump is so big, we can't easily handle well. Here, we provide 10 pre trained word2vec models with different seeds.</p> <p> </p> <p>More details about how to use it, please see paper </p>

opencc-by-4.0Mar 2017View details →
zenodo36/100

SynProtX: A Large-Scale Proteomics-Based Deep Learning Model for Predicting Synergistic Anticancer Drug Combinations

<h2>SynProtX: A Large-Scale Proteomics-Based Deep Learning Model for Predicting Synergistic Anticancer Drug Combinations</h2> <p>SynProtX is a deep learning model that integrates large-scale proteomics data, molecular graphs, and chemical fingerprints to predict synergistic effects of anticancer drug combinations. It provides robust performance across tissue-specific and study-specific datasets, enhancing reproducibility and biological relevance in drug synergy prediction.</p> <p>This Zenodo repository includes a <code>.tar.gz</code> archive containing all essential components to reproduce the experiments described in the study. This archive is designed to work seamlessly with the coding pipeline available at: <a href="https://github.com/manbaritone/SynProtX" target="_blank" rel="noopener">https://github.com/manbaritone/SynProtX</a>.</p> <h3>License:</h3> <p>Creative Commons Zero v1.0 Universal (CC0)<br>This work is released under CC0, dedicating it to the public domain. You are free to use, modify, and distribute it without restriction.</p> <h3>Archive Contents:</h3> <p>This compressed file includes:</p> <ul> <li>Datasets<br>- Tissue Datasets: <code>ALMANAC-Breast</code>, <code>ALMANAC-Lung</code>, <code>ALMANAC-Ovary</code>, <code>ALMANAC-Skin</code><br>- Study Datasets: <code>FRIEDMAN</code>, <code>ONEIL</code></li> <li>Supporting Files<br>- Raw and preprocessed data<br>- Feature dictionaries<br>- Hyperparameter configurations<br>- Trained model weights</li> </ul> <h3>Folder Structure:</h3> <blockquote> <p><code>SynProtX/</code><br><code>├── data/&nbsp; &nbsp; &nbsp; &nbsp; &nbsp; &nbsp; &nbsp; &nbsp; &nbsp; &nbsp;# Raw and preprocessed data</code><br><code>│ &nbsp; ├── export/&nbsp; &nbsp; &nbsp; &nbsp; &nbsp; &nbsp; &nbsp;# Processed protein/gene expression &amp; drug combinations</code><br><code>│ &nbsp; ├── nps/ &nbsp; &nbsp; &nbsp; &nbsp; &nbsp; &nbsp; &nbsp; &nbsp;# Numpy arrays for all datasets</code><br><code>│ &nbsp; ├── nps_intersected/ &nbsp; &nbsp;# Dataset-specific numpy arrays</code><br><code>│ &nbsp; └── raw/ &nbsp; &nbsp; &nbsp; &nbsp; &nbsp; &nbsp; &nbsp; &nbsp;# Original data from DrugComb, CCLE, COSMIC, ChEMBL V31, ProCan-DepMapSanger</code><br><code>├── feature_dicts/&nbsp; &nbsp; &nbsp; &nbsp; &nbsp; # Feature dictionaries for drug combinations</code><br><code>├── hyperparams/ &nbsp; &nbsp; &nbsp; &nbsp; &nbsp; &nbsp;# Hyperparameter configs for SynProtX-GATFP</code><br><code>│ &nbsp; ├── classification/&nbsp; &nbsp; &nbsp;# For classification tasks</code><br><code>│ &nbsp; └── regression/&nbsp; &nbsp; &nbsp; &nbsp; &nbsp;# For regression tasks</code><br><code>├── state_dict/&nbsp; &nbsp; &nbsp; &nbsp; &nbsp; &nbsp; &nbsp;# Trained model weights</code><br><code>│ &nbsp; ├── classification/&nbsp; &nbsp; &nbsp;# PyTorch checkpoints for classification</code><br><code>│ &nbsp; └── regression/&nbsp; &nbsp; &nbsp; &nbsp; &nbsp;# PyTorch checkpoints for regression</code><br><code>└── README_Zenodo.md &nbsp; &nbsp; &nbsp; &nbsp;# This file</code></p> </blockquote> <h3>For more information, please visit:</h3> <p><strong>GitHub:</strong>&nbsp;<a href="https://github.com/manbaritone/SynProtX" target="_blank" rel="noopener">https://github.com/manbaritone/SynProtX</a></p>

opencc-zeroNov 2024View details →
zenodo36/100

DeepAnnotation: A novel interpretable deep learning-based genomic selection model that integrates comprehensive functional annotations

<p>1. Update package, example dataset, and demo code of DeepAnnotation</p> <p>2. Update the transformed genotype data, the phenotype data, the comprehensive functional annotation data for Duroc prepared by RNAfold, DeepSEA, easyMF models, and the four types of input data for training DeepAnnotation model</p> <p>3. Add the conserved functional annotation</p> <p>&nbsp;</p>

opencc-zeroNov 2023View details →

ScienceDex guides

Understand access before you commit

These curated guides explain access requirements, typical timelines, costs, and reuse considerations for widely used research datasets.

Compare curated datasets

Allen Brain Atlas

Allen Brain Atlas is an Allen Institute collection of brain map atlases, datasets, APIs, and analysis tools covering mouse, human, and non-human primate brain resources.

allen-brain-atlas
neuroscienceopenDocumentation, web resources, and API references are available online.
Last verified 2026-04-30Open record

Annotated Behaviour and Observability Dataset (ABODe)

ABODe is a University of Edinburgh DataShare dataset for behavior classification in group-housed mice using home-cage video, identities, bounding boxes, ground-plate positions, and annotator labels.

abode-home-cage
behavioral-neuroscienceopenThe DataShare record exposes download links for annotations, documentation, license text, and the zipped per-snippet data directory.
Last verified 2026-04-30Open record

DANDI Archive for NWB datasets

DANDI is a BRAIN Initiative archive for publishing and sharing neurophysiology data, including electrophysiology, optophysiology, and behavioral data packaged as NWB and related standards.

dandi-nwb
electrophysiologyopenPublished Dandiset metadata and archive endpoints are available through the production DANDI API.
Last verified 2026-04-30Open record

International Brain Laboratory public data

The International Brain Laboratory public data releases expose standardized mouse decision-making experiments, including Neuropixels recordings, widefield calcium imaging, behavior, and session metadata accessed through the ONE API.

ibl
behavioral-neuroscienceopenPublic sessions can be searched and loaded from the IBL public data server through ONE.
Last verified 2026-04-29Open record

OpenNeuro

OpenNeuro is a free, open platform for sharing neuroimaging datasets, with public search, dataset pages, and download paths for web, S3, DataLad, and the OpenNeuro CLI.

openneuro
neuroscienceopenPublished datasets are available on demand over the internet.
Last verified 2026-04-29Open record