Find research datasets worth reusing
Search datasets from major research repositories and use ShareScore to quickly assess how well each record supports discovery, access, and reuse.
747
datasets available to search
ShareScore release 0.7.1
Dataset results
747 results for “Open Data”
Data from Open Data RDW: Gekentekende_voertuigen / Gekentekende_voertuigen_brandstof
<p>Data collected from the Open Data RDW repository. Data export is very slow.</p> <p> </p> <p>Web links for the two datasets:</p> <ul> <li><a href="https://opendata.rdw.nl/Voertuigen/Open-Data-RDW-Gekentekende_voertuigen/m9d7-ebf2/about_data">Open Data RDW: Gekentekende_voertuigen | Open Data | RDW</a></li> <li><a href="https://opendata.rdw.nl/Voertuigen/Open-Data-RDW-Gekentekende_voertuigen_brandstof/8ys7-d773/data_preview">Open Data RDW: Gekentekende_voertuigen_brandstof | Open Data | RDW</a></li> </ul>
Data underlying the research paper "Articulating Social Issues with Open Data: Exploring a Game Jam Approach"
<p>Contains research data underlying the following research paper:</p> <blockquote> <p>Davide Di Staso, Lærke Christiansen, Fernando Kleiman, and Marijn Janssen. 2024. Articulating Social Issues with Open Data: Exploring a Game Jam Approach. In Proceedings of the 8th International Conference on Game Jams, Hackathons and Game Creation Events (ICGJ ’24), October 11, 2024, Copenhagen, Denmark. ACM, New York, NY, USA, 7 pages. https://doi.org/10.1145/3697789.3697798</p> </blockquote> <p>The authors acknowledge the financial support from the European Union’s Horizon 2020 research and innovation program under the Marie Skłodowska-Curie grant agreement No. 955569, "Towards a sustainable Open Data ECOsystem" (ODECO).</p>
Open data on organizational success indicators
<p>This is the open data set that undergirds the analysis for "The Many Indicators of Nonprofit Success as Seen by Nonprofit Leaders"; detailed reference to be announced.</p>
Open database on traits and phylogenetic data on European pollinators
<p>(subtitle) data on wild bees, hoverflies and butterflies</p> <p>(abstract) This dataset was produced in the framework of the work package 1 (task 2) of the Horizon EU project Safeguard. The dataset encompassess trait data for three groups of pollinators: wild bees, hoverflies and butterflies, as well as phylogenetic data extracted from the GenBank. Information on traits comes from different literature sources, as well as direct measurements in collections and published databases. Some of the traits are common accross different species groups, while some traits are unique for particular group.</p> <p>(method) Litherature search and direct measurments on specimens.</p> <p>(dataset) This dataset compiles trait data for European wild bee, hoveverfly and butterfly species, and includes different categorical or numerical values for particular traits. It also contains phylogenetic information for species.</p>
An Online Integrated Development Environment for Automated Programming Assessment Systems Open Source Data
<p>This dataset accompanies the paper <em>"An Online Integrated Development Environment for Automated Programming Assessment Systems"</em>. It contains data from the usability evaluation of a feature-rich online IDE designed for integration into Automated Programming Assessment Systems (APASs). The dataset includes survey responses from 27 participants based on the Technology Acceptance Model (TAM), performance metrics such as memory usage, and qualitative user feedback. The study highlights challenges in integrating online IDEs with APASs, such as memory efficiency, load balancing, and user experience. The dataset supports further research in developing scalable, effective, and user-friendly programming education tools.<br><br>Here you can find the code changes required for the online IDE in Artemis: <a href="https://github.com/ls1intum/Artemis/pull/6706/files" target="_blank" rel="noopener">Github</a></p>
Data for From principles to practices: Open Science at European Universities. 2020-2021 EUA Open Science Survey Results
<p>This database refers to the data collected by the European University Association (EUA) for its 2020-2021 EUA Open Science Survey, which gathered responses from universities and higher education institutions across Europe. The full report published by the association is available at <a href="https://www.eua.eu/resources/publications/976:from-principles-to-practices-open-science-at-europe%E2%80%99s-universities-2020-2021-eua-open-science-survey-results.html">https://www.eua.eu/resources/publications/976:from-principles-to-practices-open-science-at-europe%E2%80%99s-universities-2020-2021-eua-open-science-survey-results.html</a> (<a href="http://doi.org/10.5281/zenodo.5062982">http://doi.org/10.5281/zenodo.5062982</a>).</p> <p>All information that could lead to the identification of individual universities and higher education institutions was removed from the database. The following files are available:</p> <ul> <li>2020-2021 EUA Open Science Survey</li> <li>Database in the following formats: .xlsx (Microsoft Excel) and .sav (IBM SPSS)</li> <li>Survey Codebook: includes information on all the variables and their coding (.xlsx)</li> <li>Data Management Plan.</li> </ul>
(open) data literacy as barrier and enabler of open government data enhancement. A systematic review of the literature.
<p>This systematic review of the literatu was conducted with the PRISMA method, to explore the contexts in which the use of open government data germinates, identifying barriers to its use and identifying, the role of data literacy among those barriers to use; and the role of open data in promoting informal learning that supports the development of critical data literacy. This file includes a codebook of the main characteristics that were studied in a systematic literature review, where data from 66 articles related to Open Data Usage were identified and coded. Also, the file includes an analysis of Cohen's Kappa, a concordance statistic used to measure the level of agreement among researchers in classifying articles on the characteristics defined in the Codebook. Finally, it includes main tables of the results' analysis.</p>
Text-fig. 5. Macroevolutionary trends related to the IC model in the first three teeth of the six families of extinct sloths, as well as specimens of the "basal Megatherioidea", Pseudoglyptodon, and Bradypus. Dashed line (- -) shows the regression including all data; solid line shows the regression after the exclusion of Octodontotherium (shown in the plot as a filled triangle). in Unexpected Inhibitory Cascade In The Molariforms Of Sloths (Folivora, Xenarthra): A Case Study In Xenarthrans Honouring Gerhard Storch'S Open-Mindedness
Text-fig. 5. Macroevolutionary trends related to the IC model in the first three teeth of the six families of extinct sloths, as well as specimens of the "basal Megatherioidea", Pseudoglyptodon, and Bradypus. Dashed line (- -) shows the regression including all data; solid line shows the regression after the exclusion of Octodontotherium (shown in the plot as a filled triangle).
Text-fig. 4. Macroevolutionary trends related to the IC model in the last three teeth of the six families of extinct sloths, as well as specimens of the "basal Megatherioidea", Pseudoglyptodon, and Bradypus. Dash-dot line (-.-) shows the regression including all data; solid line shows the regression after the exclusion of Octodontotherium (shown in the plot as a filled triangle). in Unexpected Inhibitory Cascade In The Molariforms Of Sloths (Folivora, Xenarthra): A Case Study In Xenarthrans Honouring Gerhard Storch'S Open-Mindedness
Text-fig. 4. Macroevolutionary trends related to the IC model in the last three teeth of the six families of extinct sloths, as well as specimens of the "basal Megatherioidea", Pseudoglyptodon, and Bradypus. Dash-dot line (-.-) shows the regression including all data; solid line shows the regression after the exclusion of Octodontotherium (shown in the plot as a filled triangle).
Monitoring open access publishing of NWO funded research (data)
<p>This is the dataset underlying the report "Monitoring open access publishing of NWO funded research" (https://doi.org/10.5281/zenodo.5055609). </p> <p>The report presents statistics on the extent to which publications from the period 2015–2020 funded by NWO are available in Open Access . The analyses presented in this report also cover publications funded by the Netherlands Organisation for Health Research and Development ZonMw.</p> <p>This report builds on an <a href="https://doi.org/10.5281/zenodo.4446042">earlier report</a> published in 2020 covering publications from the period 2015–2018.</p> <p> </p>
2021 UN Open GIS Challenge 1 - Training on Satellite Data Analysis and Machine Learning with QGIS (Satellite_QGIS)
<p>This dataset is part of the <a href="https://www.osgeo.org/foundation-news/2021-osgeo-un-committee-educational-challenge/?fbclid=IwAR0UvwkPO2pay7C0tJawb63eewjBGfeL9TIQpYUFccza9OIo6HAolmHXLWE">2021 UN Open GIS Challenge 1 - Training on Satellite Data Analysis and Machine Learning with QGIS (Satellite_QGIS)</a>,</p> <p>Exercise 1: Supervised Change Detection: Monitoring deglaciation in Huascaran, Peru.</p>
Open data source for "Optically reconfigurable quasi-phase-matching in silicon nitride microresonators"
<p>The folder includes includes the raw data as well as codes that were used for generation of all Figures in the paper "Optically reconfigurable quasi-phase-matching in silicon nitride microresonators".</p>
Code & Data from: Development of a low cost open-source ultrasonic device for plant height measurements
<p>We here provide code and data for the study "Development of a low cost open-source ultrasonic device for plant height measurements"</p> <p>Code:<br> - Arduino code (management of the electronic circuit): "Arduino_ultrasonic_sensor.ino"<br> - OpenSCAD code (3D-printing): "3DShells_ultrasonic_sensor.scad"<br> - R code (statistical analysis of field test): "Statistical_analysis.R"</p> <p>Data:<br> - "manual_vs_sensor_controlled.csv": this file contains the comparison between the ultrasonic device and the ruler in standardized laboratory conditions. It has three columns: "manual_value", the height value measured manually; "sensor_value", the height value obtained from the ultrasonic device; "height_range", the interval to which the height value belongs (we worked with 25 cm intervals).<br> - "manual_vs_ruler_field.csv": this file contains the comparison between the ultrasonic device and the ruler in field conditions. Plant height measurements were performed on 26 sorghum genotypes. The file has four columns: "Genotype", the id of the measured genotype; "rep" the replicate (3 plants were measured for each genotype); "manual_value", the height value measured manually; "sensor_value", the height value obtained from the ultrasonic device. When using the ruler, the operator spent 15 min and 23 s to complete all measurements in the field, and 3 min and 27 s to enter all data manually in a digital file. When using the sensor, the operator spent 10 min and 52 s to complete all measurements in the field, and manual transcription was not needed since all measurements are instantaneously saved on an SD card.</p> <p>More details on the experimental data can be found in the article "Development of a low cost open-source ultrasonic device for plant height measurements".</p> <p>We also provide a tutorial to explain how to build the ultrasonic-sensor ("tutorial.docx")</p>
CaImAn: An open source tool for scalable Calcium Imaging data Analysis
<p>Advances in fluorescence microscopy enable monitoring larger brain areas <em>in-vivo</em> with finer time resolution. The resulting data rates require reproducible analysis pipelines that are reliable, fully automated, and scalable to datasets generated over the course of months. We present CaImAn, an open-source library for calcium imaging data analysis. CaImAn provides automatic and scalable methods to address problems common to preprocessing, including motion correction, neural activity identification, and registration across different sessions of data collection. It does this while requiring minimal user intervention, with good scalability on computers ranging from laptops to high-performance computing clusters. CaImAn is suitable for two-photon and one-photon imaging, and also enables real-time analysis on streaming data.</p> <p>To benchmark the performance of CaImAn we collected and combined a corpus of manual annotations from multiple labelers on nine mouse two-photon datasets, that are contained in this open access repository. We demonstrate that CaImAn achieves near-human performance in detecting locations of active neurons.</p> <p>In order to reproduce the results of the paper or download the annotations and the raw movies, please refer to the readme.md at:</p> <p>https://github.com/flatironinstitute/CaImAn/blob/master/use_cases/eLife_scripts/README.md</p> <p> </p>
Dataset for "Open access books through open data sources: Assessing prevalence, providers, and preservation"
<p>This dataset contains the raw collected data reported on in the manuscript titled "Open access books through open data sources: Assessing prevalence, providers, and preservation" which is available here: https://doi.org/10.5281/zenodo.7305490</p> <p>One file contains the results of the digital object identifier queries, and the other data on which publication records were found to be included in which of the studied bibliometric data sources, and preservation services.</p> <p>The author is grateful to Alicia Wise and Ronald Snijder for assisting in the identification of available datasets and valuable feedback throughout the study.</p> <p>This research was commissioned by CLOCKSS, DOAB, and OAPEN.</p>
Data and Code for "Value dissonance in research(er) assessment: Individual and institutional priorities in review, promotion and tenure criteria related to research quality, quantity, openness and responsibility"
<p>This snapshot contains code and data for the preprint "Value dissonance in research(er) assessment: Individual and institutional priorities in review, promotion and tenure criteria related to research quality, quantity, openness and responsibility".</p> <p>Instructions on re-using the data and running the code can be found in the README.md.</p> <p>Changes:</p> <ul> <li>Added survey instrument and informed consent.</li> </ul>
Data and Code Supplement to: "Processes of change in a randomized clinical trial of Radically Open Dialectical Behavior Therapy (RO DBT) for adults with treatment refractory depression"
<p>Dataset to support secondary analyses reported in "Processes of change in a randomized clinical trial of Radically Open Dialectical Behavior Therapy (RO DBT) for adults with treatment refractory depression" in the Journal of Consulting and Clinical Psychology</p>
The benefit of augmenting open data with clinical data-warehouse EHR for forecasting SARS-CoV-2 hospitalizations in Bordeaux area, France
<p><strong>Objective</strong></p> <p>The aim of this study was to develop an accurate regional forecast algorithm to predict the number of hospitalized patients and to assess the benefit of the Electronic Health Records (EHR) information to perform those predictions. Materials and Methods Aggregated data from SARS-CoV-2 and weather public database and data warehouse of the Bordeaux hospital were extracted from May 16, 2020, to January 17, 2022. The outcomes were the number of hospitalized patients in the Bordeaux Hospital at 7 and 14 days. We compared the performance of different data sources, feature engineering, and machine learning models.</p> <p><strong>Results </strong></p> <p>During the period of 88 weeks, 2561 hospitalizations due to COVID-19 were recorded at the Bordeaux Hospital. The model achieving the best performance was an elastic-net penalized linear regression using all available data with a median relative error at 7 and 14 days of 0.136 [0.063; 0.223] and 0.198 [0.105; 0.302] hospitalizations, respectively. Electronic health records (EHRs) from the hospital data warehouse improved median relative error at 7 and 14 days by 10.9% and 19.8%, respectively. Graphical evaluation showed remaining forecast error was mainly due to delay in slope shift detection.</p> <p><strong>Discussion </strong></p> <p>Forecast models showed overall good performance both at 7 and 14 days which was improved by the addition of the data from Bordeaux Hospital data warehouse.</p> <p><strong>Conclusions </strong></p> <p>The development of hospital data warehouses might help to get more specific and faster information than traditional surveillance systems, which in turn will help to improve epidemic forecasting at a larger and finer scale.</p>
WorldCereal open global harmonized reference data repository (CC-BY-SA licensed data sets)
<p>Within the<strong> ESA funded</strong> WorldCereal project we have built an open harmonized reference data repository at global extent for model training or product validation in support of land cover and crop type mapping. Data from 2017 onwards were collected from many different sources and then harmonized, annotated and evaluated. These steps are explained in the harmonization protocol (10.5281/zenodo.7584463). This protocol also clarifies the naming convention of the shape files and the WorldCereal attributes (LC, CT, IRR, valtime and sampleID) that were added to the original data sets.</p> <p>This publication includes those harmonized data sets of which the original data set was published under the CC-BY-SA license or a license similar to CC-BY-SA. See document "_In-situ-data-World-Cereal - license - CC-BY-SA.pdf" for an overview of the original data sets.</p>
WorldCereal open global harmonized reference data repository (CC-BY licensed data sets)
<p>Within the <strong>ESA funded </strong>WorldCereal project we have built an open harmonized reference data repository at global extent for model training or product validation in support of land cover and crop type mapping. Data from 2017 onwards were collected from many different sources and then harmonized, annotated and evaluated. These steps are explained in the harmonization protocol (10.5281/zenodo.7584463). This protocol also clarifies the naming convention of the shape files and the WorldCereal attributes (LC, CT, IRR, valtime and sampleID) that were added to the original data sets.</p> <p>This publication includes those harmonized data sets of which the original data set was published under the CC-BY license or a license similar to CC-BY. See document "_In-situ-data-World-Cereal - license - CC-BY.pdf" for an overview of the original data sets. </p>
ScienceDex guides
Understand access before you commit
These curated guides explain access requirements, typical timelines, costs, and reuse considerations for widely used research datasets.
Allen Brain Atlas
Allen Brain Atlas is an Allen Institute collection of brain map atlases, datasets, APIs, and analysis tools covering mouse, human, and non-human primate brain resources.
Annotated Behaviour and Observability Dataset (ABODe)
ABODe is a University of Edinburgh DataShare dataset for behavior classification in group-housed mice using home-cage video, identities, bounding boxes, ground-plate positions, and annotator labels.
DANDI Archive for NWB datasets
DANDI is a BRAIN Initiative archive for publishing and sharing neurophysiology data, including electrophysiology, optophysiology, and behavioral data packaged as NWB and related standards.
International Brain Laboratory public data
The International Brain Laboratory public data releases expose standardized mouse decision-making experiments, including Neuropixels recordings, widefield calcium imaging, behavior, and session metadata accessed through the ONE API.
OpenNeuro
OpenNeuro is a free, open platform for sharing neuroimaging datasets, with public search, dataset pages, and download paths for web, S3, DataLad, and the OpenNeuro CLI.